AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
693 results
Model Releases

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds resea…

DGX agent

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds research-backed models autonomously. Pushed a benchmark from 10%

model-releasesclem-delangue--x
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

I think the CoALA paper's classification system of semantic/episodic/procedural is maybe the closest thing we have to a standard for agent m…

DGX agent

The CoALA paper proposes a classification system for agent memory that distinguishes between semantic memory (facts and concepts), episodic memory (specific experiences and events), and procedural mem

agentsharrison-chase--x
21 Apr 2026
Model Releases

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Co…

DGX agent

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Codex land near the human median, but with far tighter dispers

model-releasesethan-mollick--x
20 Apr 2026
Industry

First paper, microsite & NotebookLM https://ii.inc/web/releases/one-postulate

DGX agent

Emad Mostaque announced the release of Stability AI's first paper, microsite, and integration with NotebookLM, likely detailing new research findings or technical capabilities. The announcement appear

industryemad-mostaque--x
20 Apr 2026
Model Releases

The Top AI Papers of the Week (April 13 - 19) - AlphaEval - AiScientist - Auto-Diagnose - Nemotron 3 Super - Subliminal Learning - Automated…

DGX agent

The Top AI Papers of the Week (April 13 - 19) - AlphaEval - AiScientist - Auto-Diagnose - Nemotron 3 Super - Subliminal Learning - Automated W2S Researcher - Memory Transfer Learning Read on for more:

model-releasesdair-ai--x
19 Apr 2026
Model Releases

A lot of papers coming out are still focused on GPT-4, but you could extrapolate their effects to GPT-5, etc. Much harder to know what the i…

DGX agent

A lot of papers coming out are still focused on GPT-4, but you could extrapolate their effects to GPT-5, etc. Much harder to know what the impacts of Claude Code/Codex etc. are because they are so new

model-releasesethan-mollick--x
16 Apr 2026
Industry

Research paper link: https://www.pnas.org/doi/10.1073/pnas.2519129123 I break down stories like this every day in my free newsletter. Keep u…

DGX agent

Research paper link: https://www.pnas.org/doi/10.1073/pnas.2519129123 I break down stories like this every day in my free newsletter. Keep up with the latest in AI/Robotics in 5 min a day: https://www

industryrowan-cheung--x
14 Apr 2026
Industry

Research paper link: https://www.cell.com/cell/fulltext/S0092-8674%2825%2901312-1 I break down stories like this every day in my free newsle…

DGX agent

Research paper link: https://www.cell.com/cell/fulltext/S0092-8674%2825%2901312-1 I break down stories like this every day in my free newsletter. Keep up with the latest in AI/Robotics in 5 min a day:

industryrowan-cheung--x
13 Apr 2026
Agents

The Top AI Papers of the Week (April 6 - 12) - Memento - Neural Computers - The Universal Verifier - Agent Skills in the Wild - Memory Intel…

DGX agent

The Top AI Papers of the Week (April 6 - 12) - Memento - Neural Computers - The Universal Verifier - Agent Skills in the Wild - Memory Intelligence Agent (MIA) - Single-Agent vs Multi-Agent LLMs - Sca

agentsdair-ai--x
12 Apr 2026
Model Releases

I’ve uploaded a new paper on arXiv (co-authored by @rasbt): MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning In Parameter-Efficient…

DGX agent

I’ve uploaded a new paper on arXiv (co-authored by @rasbt): MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning In Parameter-Efficient Fine-Tuning, a key question may not just be how low-rank th

model-releasessebastian-raschka--x
8 Apr 2026
Applications

Some approaches: 1) Multiple reviews. Some papers already show having many AI team members review a problem reduces errors 2) Building in te…

DGX agent

Some approaches: 1) Multiple reviews. Some papers already show having many AI team members review a problem reduces errors 2) Building in tests and checkpoints 3) Multiple independent answers that are

applicationsethan-mollick--x
8 Apr 2026
Model Releases

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimo…

DGX agent

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimodal generative modeling could accelerate protein sciences! ⬇

model-releasesyoshua-bengio--x
8 Apr 2026
Research

Top Community Contributors: @SHL0MS (7 PRs) — p5js creative coding skill, manim-video skill + 5 reference expansions, research-paper-writing…

DGX agent

Top Community Contributors: @SHL0MS (7 PRs) — p5js creative coding skill, manim-video skill + 5 reference expansions, research-paper-writing, Nous OAuth fix, manim fix @sidbing (3 PRs) — Firecrawl clo

researchnous-research--x
8 Apr 2026
Applications

More on the @nytimes piece about that $1.8B, two-person, AI company ... not the paper's finest moment in quick retrospect. And in the health…

DGX agent

More on the @nytimes piece about that $1.8B, two-person, AI company ... not the paper's finest moment in quick retrospect. And in the healthcare context no less. Our friend @GaryMarcus was on this a f

applicationsgary-marcus--x
7 Apr 2026
Tools

4/ Escaping the Verifier: Learning to Reason via Demonstrations (RARO) Paper: https://arxiv.org/abs/2511.21667

DGX agent

RARO (Reasoning via Demonstrations) is a method for training AI models to improve reasoning capabilities by learning from demonstrations rather than relying solely on external verifiers. The approach

toolstogether-ai--x
1 Jul 2026
Safety

receipts for most points can found here, if you read this paper closely: https://nautil.us/deep-learning-is-hitting-a-wall-238440

DGX agent

Gary Marcus references a Nautilus article arguing that deep learning is encountering fundamental limitations, suggesting readers can find supporting evidence and detailed arguments for this perspectiv

safetygary-marcus--x
23 May 2026
Research

Paper: https://arxiv.org/abs/2605.06554 Code: https://github.com/ighoshsubho/lighthouse-attention HF: https://huggingface.co/papers/2605.065…

DGX agent

Lighthouse Attention is a novel attention mechanism that improves efficiency in transformer models by selectively focusing computation on the most relevant tokens, similar to how a lighthouse beam ill

researchnous-research--x
15 May 2026
Agents

// Agentic World Modeling // Massive 40-author survey just dropped. Cleanest taxonomy of world models in agent research I've seen. (bookmark…

DGX agent

// Agentic World Modeling // Massive 40-author survey just dropped. Cleanest taxonomy of world models in agent research I've seen. (bookmark it) The paper proposes a 'levels × laws' framework. Three c

agentsdair-ai--x
27 Apr 2026
Tools

1/ DSGym: A Holistic Framework for Evaluating and Training Data Science Agents Paper: https://arxiv.org/abs/2601.16344

DGX agent

DSGym is a comprehensive framework designed to evaluate and train AI agents for data science tasks, providing a structured environment for benchmarking agent performance across various data science wo

toolstogether-ai--x
1 Jul 2026
Agents

2/ ThunderAgent: A Simple, Fast and Program-Aware Agentic Inference System Paper: https://arxiv.org/abs/2602.13692

DGX agent

ThunderAgent is an agentic inference system designed for fast and efficient execution of AI agent programs, developed by Together AI. The system appears to optimize program-aware inference by leveragi

agentstogether-ai--x
1 Jul 2026
Tools

3/ Learning to Discover at Test Time (TTT-Discover) Paper: https://arxiv.org/abs/2601.16175

DGX agent

TTT-Discover is a method that enables models to learn and discover patterns during test time rather than only during training, allowing for adaptation to new data distributions at inference. The appro

toolstogether-ai--x
1 Jul 2026
Tools

6/ When RL Meets Adaptive Speculative Training: A Unified Training-Serving System (Aurora) Paper: https://arxiv.org/abs/2602.06932

DGX agent

Aurora is a unified training-serving system that integrates reinforcement learning with adaptive speculative training to optimize large language model inference and training efficiency. The system dyn

toolstogether-ai--x
1 Jul 2026
Tools

7/ Untied Ulysses: Memory-Efficient Context Parallelism via Headwise Chunking Paper: https://arxiv.org/abs/2602.21196

DGX agent

Untied Ulysses is a memory-efficient technique for context parallelism that processes attention heads in chunks rather than sequences, reducing memory overhead during transformer inference and trainin

toolstogether-ai--x
1 Jul 2026
Tools

8/ Opportunistic Expert Activation: Batch-Aware Expert Routing for Faster Decode Without Retraining (OEA) Paper: https://arxiv.org/abs/2511.…

DGX agent

Opportunistic Expert Activation (OEA) is a batch-aware expert routing technique for mixture-of-experts models that enables faster decoding without requiring model retraining. The method optimizes whic

toolstogether-ai--x
1 Jul 2026
Hardware

9/ ParallelKernelBench: Benchmarking LLMs on Multi-GPU Kernel Generation Paper: https://www.alphaxiv.org/abs/2606.parallel-kernel-bench

DGX agent

ParallelKernelBench is a benchmarking framework designed to evaluate large language models' ability to generate optimized GPU kernels for multi-GPU computing environments. The benchmark assesses LLMs

hardwaretogether-ai--x
1 Jul 2026
Local Ai

Read the technical paper on Krea 2 https://www.krea.ai/blog/krea-2-technical-report Download the model weights https://github.com/krea-ai/kr…

DGX agent

Krea 2 is a technical advancement in AI image generation with newly released model weights available for download on GitHub. The technical report details the improvements and capabilities of this vers

local-aicomfyui--x
23 Jun 2026
Safety

a shout out to the paper: https://arxiv.org/html/2605.25376v1 'KYA: A Framework-Agnostic Trust Layer for Autonomous Systems with Verifiable …

DGX agent

KYA is a framework-agnostic trust layer designed for autonomous systems that provides verifiable guarantees, addressing the need for trustworthy and transparent operation of AI agents across different

safetyyohei-nakajima--x
26 May 2026
Research

Learning to Orchestrate Agents in Natural Language with the Conductor Fugu Blog: https://sakana.ai/fugu-beta Paper: https://arxiv.org/abs/25…

DGX agent

Conductor is a method for orchestrating multiple AI agents through natural language instructions, enabling coordinated multi-agent systems where a central 'conductor' agent directs specialized agents

researchdavid-ha--x
28 Apr 2026
Model Releases

ChatGPT diagnosed 40 million people with a disease that was invented as a joke. Not a real disease. Not a misunderstood disease. A completel…

DGX agent

ChatGPT diagnosed 40 million people with a disease that was invented as a joke. Not a real disease. Not a misunderstood disease. A completely fictional condition with a fake name, fake papers, and fak

model-releasesgary-marcus--x
29 May 2026
Industry

ANNOUNCEMENT: WE’RE SAVING SCIENCE! We’re often told that science is “self-correcting.” But that’s not really true. Science doesn’t correct …

DGX agent

ANNOUNCEMENT: WE’RE SAVING SCIENCE! We’re often told that science is “self-correcting.” But that’s not really true. Science doesn’t correct itself like a thermostat adjusting the temperature in your h

industryelon-musk--x
27 May 2026
Agents

Very cool idea to convert memory to skills. (bookmark it) Most agent memory systems retrieve past traces as passive context. MSCE turns them…

DGX agent

Very cool idea to convert memory to skills. (bookmark it) Most agent memory systems retrieve past traces as passive context. MSCE turns them into executable skills instead. The training-free framework

agentsdair-ai--x
21 Jul 2026
Agents

Pay attention to this one, AI devs. If you're building multi-agent systems, you're probably wiring static org charts. New research argues th…

DGX agent

Pay attention to this one, AI devs. If you're building multi-agent systems, you're probably wiring static org charts. New research argues they should look more like a labor market. The paper introduce

agentsdair-ai--x
27 Apr 2026
Model Releases

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our…

DGX agent

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our new paper: “TRINITY: An Evolved LLM Coordinator”, published

model-releasesdavid-ha--x
25 Apr 2026
Model Releases

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, …

DGX agent

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, 17 real tasks from 3 repositories, with context-injection st

model-releasesdair-ai--x
1 Aug 2026
Agents

my weekend hobby: self improvement research

DGX agent

my weekend hobby: self improvement research in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: An Auditable, He

agentsyohei-nakajima--x
10 Jun 2026
Agents

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture'

DGX agent

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture' babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is t

agentsyohei-nakajima--x
28 May 2026
Agents

recommended reading.

DGX agent

recommended reading. babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-Sourced Reactive Graphs for Auditable, Forkable Agentic Systems

agentsyohei-nakajima--x
22 May 2026
Safety

A true exponential!

DGX agent

A true exponential! Oy. According to a new paper in The Lancet, the rate of made-up citations in biomedical papers has increased by more than 12x since 2023. https://www.thelancet.com/journals/lancet/

safetygary-marcus--x
12 May 2026
Industry

Oh?

DGX agent

Oh? Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One characteristic of the algebra A positive cosmological cons

industryemad-mostaque--x
23 Apr 2026
Hardware

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind…

DGX agent

This is crazy. ml-intern just passed the @huggingface internship test in 15 minutes. The task: replicate a research baseline from a DeepMind paper on test-time compute scaling. Here's what the agent d

hardwareclem-delangue--x
23 Apr 2026
Industry

Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One ch…

DGX agent

Today we release a novel AI-assisted resolution of one of physics’ longest-standing questions. Given only: • Relativity as an axiom • One characteristic of the algebra A positive cosmological constant

industryemad-mostaque--x
22 Apr 2026
Industry

What's cooler than finding a 27-year-old bug in OpenBSD? Finding a positive cosmological constant hiding for over a century in the algebra o…

DGX agent

What's cooler than finding a 27-year-old bug in OpenBSD? Finding a positive cosmological constant hiding for over a century in the algebra of relativity🌌 No new physics or math needed🧮 Possibly the mo

industryemad-mostaque--x
22 Apr 2026
Model Releases

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real rese…

DGX agent

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real research loop that our ML researchers do every day. You give it

model-releasesclem-delangue--x
21 Apr 2026
Model Releases

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research…

DGX agent

// The Bitter Lesson of Tool Calling // Tool calling is a design choice, and the defaults are quietly costing accuracy. How so? New research releases a generation-spanning comparison of programmatic t

model-releasesdair-ai--x
10 Aug 2026
Model Releases

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. C…

DGX agent

Skill libraries are shipping in agent harnesses on the assumption that writing skills down compounds. A new benchmark tests that directly. ContinualSkillBench covers five domains, each with 100 interc

model-releasesdair-ai--x
5 Aug 2026
Model Releases

New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate…

DGX agent

New research from Meta and CMU. This one is on agentic context management for long horizon tasks. (bookmark it) Production agents accumulate context every turn. The usual fix compresses on a token thr

model-releasesdair-ai--x
28 Jul 2026
Model Releases

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of A…

DGX agent

LLM Wikis are being slept on. I argue that creating knowledge bases with LLMs or coding agents is one of the most valuable applications of AI today. It's about being intentional in building and scalin

model-releasesdair-ai--x
2 Jul 2026
Tutorials

🚨 In a surprise to NOBODY, a study shows that generative AI use raises homework scores, but substantially reduces learning. My takeaways fo…

DGX agent

🚨 In a surprise to NOBODY, a study shows that generative AI use raises homework scores, but substantially reduces learning. My takeaways for AI ethicists and educators: The study analyzed data from 26

tutorialsgary-marcus--x
25 Jun 2026
← Previous
1…45678…15
Next →