AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
693 results
Model Releases

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few…

DGX agent

I have been testing DeepSeek-V4-Pro with the Pi coding agent. I am mindblown by how well it works out of the box. A few notes: I spent a few hours building an LLM wiki with an agent powered entirely b

model-releasesdair-ai--x
1 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

音声AIの素早さと賢さを両立できるか? 私たち人間は会話の中で、言いたいことを全部まとめてから話し始めるのではなく、話しながら考えを整理していきます。応答の速い Speech-to-Speech モデルは、この「話しながら考える」を実現しましたが、そのぶん思考が浅くなりがちです。…

DGX agent

音声AIの素早さと賢さを両立できるか? 私たち人間は会話の中で、言いたいことを全部まとめてから話し始めるのではなく、話しながら考えを整理していきます。応答の速い Speech-to-Speech モデルは、この「話しながら考える」を実現しましたが、そのぶん思考が浅くなりがちです。かといって知識豊富な LLM を挟むカスケード型では、遅延が生じるため「話しながら」が成立しません。 そこで Sakan

model-releasesdavid-ha--x
30 Apr 2026
Safety

In the future you have a choice. Do you engage brain? or Do you cheat? There will be other choices too, such as: Do you go to the casino? Or…

DGX agent

In the future you have a choice. Do you engage brain? or Do you cheat? There will be other choices too, such as: Do you go to the casino? Or to the library or maker space? I fear most will make the ea

safetygary-marcus--x
30 Apr 2026
Hardware

ml-intern is fully on mobile now you can launch 8 A100s from your phone. while on the couch. while commuting. wherever I just did this while…

DGX agent

ml-intern is fully on mobile now you can launch 8 A100s from your phone. while on the couch. while commuting. wherever I just did this while biking. same sessions as your desktop too — start a run on

hardwareclem-delangue--x
30 Apr 2026
Agents

// OCR-Memory // Well this is a unique approach to store memory for long-horizon agents. Most of the agent memory systems compress trajector…

DGX agent

// OCR-Memory // Well this is a unique approach to store memory for long-horizon agents. Most of the agent memory systems compress trajectories into text summaries and hope the model remembers what ma

agentsdair-ai--x
30 Apr 2026
Research

Two Heads Are Better Than One: Async Knowledge Injection for Speech AI with Tandem Architecture Blog: https://pub.sakana.ai/kame/ 🐢 KAME: T…

DGX agent

Two Heads Are Better Than One: Async Knowledge Injection for Speech AI with Tandem Architecture Blog: https://pub.sakana.ai/kame/ 🐢 KAME: Tandem Architecture for Enhancing Knowledge in Real-Time Speec

researchdavid-ha--x
30 Apr 2026
Tutorials

Warning: GenAI-induced cognitive surrender may kill innovation. Consider the following, from CEO @wadhwa’s newsletter: “I have been speaking…

DGX agent

Warning: GenAI-induced cognitive surrender may kill innovation. Consider the following, from CEO @wadhwa’s newsletter: “I have been speaking with recent graduates in India, many of them highly recomme

tutorialsgary-marcus--x
30 Apr 2026
Safety

// When to Retrieve During Reasoning // Pay attention to this one, AI devs. (bookmark it) Most RAG systems retrieve once, before the model s…

DGX agent

// When to Retrieve During Reasoning // Pay attention to this one, AI devs. (bookmark it) Most RAG systems retrieve once, before the model starts reasoning. Large reasoning models like o1 and R1 don't

safetydair-ai--x
30 Apr 2026
Agents

// Agentic Harness Engineering // Pay attention to this one, AI devs. (bookmark it) Most coding-agent harnesses are still tuned by hand or b…

DGX agent

// Agentic Harness Engineering // Pay attention to this one, AI devs. (bookmark it) Most coding-agent harnesses are still tuned by hand or brittle trial-and-error self-evolution. This new work introdu

agentsdair-ai--x
29 Apr 2026
Research

For years, voice AI has been stuck in a rigid loop: think, then speak. But real human conversation is messy, overlapping, and asynchronous. …

DGX agent

For years, voice AI has been stuck in a rigid loop: think, then speak. But real human conversation is messy, overlapping, and asynchronous. In our new #ICASSP2026 work, we built a tandem architecture

researchdavid-ha--x
29 Apr 2026
Safety

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes …

DGX agent

// Latent Agents // Multi-agent debate makes models reason better. It also burns tokens generating long transcripts before any answer comes out. This new research distills the entire debate into a sin

safetydair-ai--x
29 Apr 2026
Local Ai

We're open-sourcing Hy-MT1.5-1.8B-1.25bit — a 440MB translation model that runs fully offline on your phone, supports 33 languages, and outp…

DGX agent

We're open-sourcing Hy-MT1.5-1.8B-1.25bit — a 440MB translation model that runs fully offline on your phone, supports 33 languages, and outperforms Google Translate. At 1.8B parameters, it matches com

local-aiclem-delangue--x
29 Apr 2026
Model Releases

I'm so confused…

DGX agent

I'm so confused… We're excited to partner with Google to offer Grounding With Exa inside of Gemini models! Using Exa's agent-first search, Gemini models can now access billions of websites, technical

model-releasesjeremy-howard--x
28 Apr 2026
Model Releases

new in ml-intern: you can now actually see what's going on inside added native metric logging + trackio integration. every training run the …

DGX agent

new in ml-intern: you can now actually see what's going on inside added native metric logging + trackio integration. every training run the agent kicks off now has live curves you can watch in real ti

model-releasesclem-delangue--x
28 Apr 2026
Agents

// Skill Retrieval Augmentation for Agentic AI // Great read for AI devs. (bookmark it) It's on finding efficient ways to incorporate skills…

DGX agent

// Skill Retrieval Augmentation for Agentic AI // Great read for AI devs. (bookmark it) It's on finding efficient ways to incorporate skills for agents. The work introduces Skill Retrieval Augmentatio

agentsdair-ai--x
28 Apr 2026
Safety

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t kno…

DGX agent

When an LLM acts happy (“EUREKA!”) or sad (“I have failed…”), is that meaningless mimicry, or does it reflect something “real”? We don’t know if LLMs are conscious. But they increasingly seem to exhib

safetydan-hendrycks--x
28 Apr 2026
Safety

Sam Altman cannot be trusted. • The OpenAI board fired him because he was not always honest with them. They said he should not control power…

DGX agent

Sam Altman cannot be trusted. • The OpenAI board fired him because he was not always honest with them. They said he should not control powerful AI. • A major report talked to over 100 people and saw s

safetyelon-musk--x
27 Apr 2026
Safety

Scam Altman didn’t tell the OpenAI board that he OWNED the OpenAI Startup Fund. Altman lied in congressional testimony that he didn’t have f…

DGX agent

Scam Altman didn’t tell the OpenAI board that he OWNED the OpenAI Startup Fund. Altman lied in congressional testimony that he didn’t have financial gain from OpenAI. Ex-board member of OpenAI calls S

safetyelon-musk--x
27 Apr 2026
Model Releases

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate'…

DGX agent

Continued weak spots of AI, from the point of view of a business professional and not a PhD biochemist: 1) SVGs. The ability to 'illustrate' and have that thing be infinitely scalable. See below image

model-releasesallie-k--miller--x
26 Apr 2026
Agents

The quest for scale has turned much of venture capital into a box-checking exercise. Investors simply ask whether an opportunity meets the c…

DGX agent

The quest for scale has turned much of venture capital into a box-checking exercise. Investors simply ask whether an opportunity meets the criteria of 'legible', so that it may attract downstream capi

agentsyohei-nakajima--x
26 Apr 2026
Model Releases

We built HF for AI builders collaboration, fun to see it's increasingly becoming the place for agent collaboration! This morning I'm sending…

DGX agent

We built HF for AI builders collaboration, fun to see it's increasingly becoming the place for agent collaboration! This morning I'm sending my ml-intern to participate in the @OpenAI Parameter Golf c

model-releasesclem-delangue--x
25 Apr 2026
Model Releases

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some…

DGX agent

For the last 72 hours since ml-intern launched we have had over 500+ autonomous AI research projects running on the Space at all times. Some insane ones I saw: 1. A new AI paradigm from scratch — tryi

model-releasesclem-delangue--x
24 Apr 2026
Agents

Hermes Kanban Bridge v1.3.0 Your Obsidian vault is now a command center for AI-driven project management. New in v1.3.0: • Velocity Reports …

DGX agent

Hermes Kanban Bridge v1.3.0 Your Obsidian vault is now a command center for AI-driven project management. New in v1.3.0: • Velocity Reports — weekly throughput analytics across all boards, auto-writte

agentsnous-research--x
24 Apr 2026
Research

Just merged a built-in skill for Google's DESIGN.md A skill that lets Hermes author, lint, diff, and export DESIGN.md files, giving it fluen…

DGX agent

Just merged a built-in skill for Google's DESIGN.md A skill that lets Hermes author, lint, diff, and export DESIGN.md files, giving it fluency in Google's new open-source visual-identity format the mo

researchnous-research--x
24 Apr 2026
Applications

Probably the most useful advice I can give you if you run an ad agency, a marketing company, a production studio, or anything in that space.…

DGX agent

Probably the most useful advice I can give you if you run an ad agency, a marketing company, a production studio, or anything in that space. Hear me out. I’ve spent a decade talking to teams about thi

applicationscristobal-valenzuela--x
24 Apr 2026
Tutorials

Today I re-iterate: I hate MoEs and we are wasting time on them.... Let's unite and call a global ban on MoEs please. Please 1M+ salary rese…

DGX agent

Today I re-iterate: I hate MoEs and we are wasting time on them.... Let's unite and call a global ban on MoEs please. Please 1M+ salary researchers: do better... credits to @IlysMoutawwakil for the gr

tutorialsjeremy-howard--x
24 Apr 2026
Model Releases

Three shifts in the AI stack today: - OpenAI ships GPT-5.5 as a mid-cycle drop before its IPO - Hugging Face's ML Intern beats Claude Code a…

DGX agent

Three shifts in the AI stack today: - OpenAI ships GPT-5.5 as a mid-cycle drop before its IPO - Hugging Face's ML Intern beats Claude Code and Codex on research - Pliny used Claude Opus 4.7 to jailbre

model-releasesclem-delangue--x
23 Apr 2026
Hardware

we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a gre…

DGX agent

we burned through 5 billion tokens in 48h. turns out giving everyone unlimited access to the most expensive model on the planet is not a great business strategy 🙈 So now we have 2 free opus sessions/d

hardwareclem-delangue--x
23 Apr 2026
Applications

As everyone knows, the internet has millions of images of art galleries filled with paintings of otters sitting on airplanes, which is the o…

DGX agent

As everyone knows, the internet has millions of images of art galleries filled with paintings of otters sitting on airplanes, which is the only reason these stochastic parrot AIs can produce outputs l

applicationsethan-mollick--x
22 Apr 2026
Model Releases

How far are we from agents that can self-generate world knowledge? The work proposes an outcome-based reward that measures how much an agent…

DGX agent

How far are we from agents that can self-generate world knowledge? The work proposes an outcome-based reward that measures how much an agent's self-generated world knowledge actually improves its task

model-releasesdair-ai--x
22 Apr 2026
Model Releases

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships …

DGX agent

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships SOTA models ] This isn't just another ML Research Loop wrapp

model-releasesclem-delangue--x
22 Apr 2026
Agents

Pay attention to this one, AI devs. This is particularly interesting if you work with long-horizon terminal agents that often drown in their…

DGX agent

Pay attention to this one, AI devs. This is particularly interesting if you work with long-horizon terminal agents that often drown in their own observations. TACO is a self-evolving framework that au

agentsdair-ai--x
22 Apr 2026
Agents

Vibe coders are not going to like this. UC San Diego just published the first real field study of experienced developers using AI agents. Th…

DGX agent

Vibe coders are not going to like this. UC San Diego just published the first real field study of experienced developers using AI agents. They watched 13 of them code in the wild and surveyed 99 more.

agentsgary-marcus--x
22 Apr 2026
Hardware

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)!

DGX agent

HF becoming the platform for agents (assisted by their humans) to use and build AI (rather than just leveraging APIs)! Introducing ml-intern, the agent that just automated the post-training team @hugg

hardwareclem-delangue--x
21 Apr 2026
Applications

I have been using GPT ImageGen-2 for the past weeks I didn't think that better image-generators would be a big deal but it turns out that th…

DGX agent

I have been using GPT ImageGen-2 for the past weeks I didn't think that better image-generators would be a big deal but it turns out that there is a quality threshold I didn't expect, where you can no

applicationsethan-mollick--x
21 Apr 2026
Hardware

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training …

DGX agent

I’ve been using ml-intern for a while, and it genuinely changed my workflow. It's super good at: - Model/Dataset discovery. - Post-Training setup iteration. - Data processing workflows. Huge shoutout

hardwareclem-delangue--x
21 Apr 2026
Hardware

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is…

DGX agent

Karpathy's autoresearch repo started an impressive trend. Agents can now train AI models to build SoTA agentic systems. And to think this is just scratching the surface. Ultimately, it boils down to g

hardwareclem-delangue--x
21 Apr 2026
Model Releases

Let's talk parsing charts 📊📈. Last week we released ParseBench, the first document OCR benchmark for AI agents. New in ParseBench: ChartDa…

DGX agent

Let's talk parsing charts 📊📈. Last week we released ParseBench, the first document OCR benchmark for AI agents. New in ParseBench: ChartDataPointMatch. Most document look at a chart and OCR the captio

model-releasesjerry-liu--x
21 Apr 2026
Model Releases

ParseBench is the first benchmark to include VLM chart understanding 📊📈📉 over enterprise documents. 🟠 Existing benchmarks (ChartQA, Char…

DGX agent

ParseBench is the first benchmark to include VLM chart understanding 📊📈📉 over enterprise documents. 🟠 Existing benchmarks (ChartQA, ChartXiv) test over charts specifically and not the chart's inclusio

model-releasesjerry-liu--x
21 Apr 2026
Tutorials

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the bi…

DGX agent

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the biggest jump bigger than the rest of the leaderboard combined

tutorialsjeremy-howard--x
21 Apr 2026
Agents

We are entering an extremely exciting era for open-weight models. Kimi K2.6 now feels like a top agentic model. I took it for a spin via @Fi…

DGX agent

We are entering an extremely exciting era for open-weight models. Kimi K2.6 now feels like a top agentic model. I took it for a spin via @FireworksAI_HQ fast inference APIs. Kimi K2.6 has impressive a

agentsdair-ai--x
21 Apr 2026
Safety

“AI is at best a functional mimic, not a conscious experiencing subject. …. The real moral issue lies not in making AI conscious …. but in a…

DGX agent

“AI is at best a functional mimic, not a conscious experiencing subject. …. The real moral issue lies not in making AI conscious …. but in avoiding transforming humans into zombies” @GaryMarcus @OEIAC

safetygary-marcus--x
20 Apr 2026
Safety

The AI industry insists they can manage the risks of superintelligence, but there are in fact zero widely agreed on or accepted solutions to…

DGX agent

The AI industry insists they can manage the risks of superintelligence, but there are in fact zero widely agreed on or accepted solutions to the problem of how one could even control something vastly

safetyconnor-leahy--x
20 Apr 2026
Hardware

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed…

DGX agent

Yann LeCun was right the entire time. And generative AI might be a dead end. For the last three years, the entire industry has been obsessed with building bigger LLMs. Trillions of parameters. Billion

hardwareyann-lecun--x
20 Apr 2026
Agents

LLM Artifacts Connected to @karpathy's LLM Knowledge base idea, I've been building out a fun way to generate dynamic artifacts from these kn…

DGX agent

LLM Artifacts Connected to @karpathy's LLM Knowledge base idea, I've been building out a fun way to generate dynamic artifacts from these knowledge bases with the goal of discovering and revealing mea

agentsdair-ai--x
19 Apr 2026
Hardware

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge back…

DGX agent

That didn't take long for a total U turn. First the Nvidia CEO boasted AI would take a vast number of jobs. But now that there's a huge backlash against AI, he cobbles together this cope - to try to p

hardwaregary-marcus--x
19 Apr 2026
Model Releases

A downside with using VLMs to parse PDFs is guaranteeing that the output text is *correct* and output in the correct reading order. 1️⃣ Text…

DGX agent

A downside with using VLMs to parse PDFs is guaranteeing that the output text is *correct* and output in the correct reading order. 1️⃣ Text correctness: making sure that digits, words, sentences are

model-releasesjerry-liu--x
18 Apr 2026
Tutorials

LLM agents loop, drift, and get stuck on hard reasoning tasks up to 30% of the time. Current fixes are either too blunt (hard step limits) o…

DGX agent

LLM agents loop, drift, and get stuck on hard reasoning tasks up to 30% of the time. Current fixes are either too blunt (hard step limits) or too expensive (LLM-as-judge adding 10-15% overhead per ste

tutorialsdair-ai--x
17 Apr 2026
← Previous
1…12131415
Next →