AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,056 results
Model Releases

DeepSeek V4 paper full version is out, FP4 QAT details and stability tricks [D]

DGX agent

DeepSeek released the full technical report for DeepSeek-V4 on April 24, 2026, titled 'DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence.' The paper details FP4 quantization-awa

model-releasesr-machinelearning
9 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Tools

🥈 2nd Place — Paper Trail by Ryan Hoare The chaos of school letters, party invites, and class WhatsApp groups, turned into a calm family pl…

DGX agent

🥈 2nd Place — Paper Trail by Ryan Hoare The chaos of school letters, party invites, and class WhatsApp groups, turned into a calm family plan. What we liked: useful for the family in an obvious-once-y

toolsreplit--x
8 May 2026
Safety

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying t…

DGX agent

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to control AIs is a limited strategy, and that a stable, mutu

safetydan-hendrycks--x
7 May 2026
Model Releases

💫Very happy to release NeuralBench, to benchmark Neuro AI models and datasets in the open! 🧵Thread, 💻Code, 📝White Paper below:

DGX agent

💫Very happy to release NeuralBench, to benchmark Neuro AI models and datasets in the open! 🧵Thread, 💻Code, 📝White Paper below: 🧠 Introducing NeuralBench: a unified, open-source framework to benchmark

model-releasesyann-lecun--x
6 May 2026
Agents

The Top AI Papers of the Week (April 26 - May 3) - Latent Agents - RecursiveMAS - OneManCompany - AgenticQwen-30B-A3B - Agentic World Modeli…

DGX agent

The Top AI Papers of the Week (April 26 - May 3) - Latent Agents - RecursiveMAS - OneManCompany - AgenticQwen-30B-A3B - Agentic World Modeling - Agentic Harness Engineering - From Skill Text to Skill

agentsdair-ai--x
3 May 2026
Applications

New paper (on an old AI) tests o1 against doctors on medical benchmarks & real ER cases: “across a variety of scenarios and applications, th…

DGX agent

New paper (on an old AI) tests o1 against doctors on medical benchmarks & real ER cases: “across a variety of scenarios and applications, the large language model outperformed both human physicians an

applicationsethan-mollick--x
1 May 2026
Agents

NEW paper: Recursive Multi-Agent Systems

DGX agent

NEW paper: Recursive Multi-Agent Systems // Recursive Multi-Agent Systems // Great read for the weekend. (bookmark it) Multi-agent systems often pass full text messages between agents at every step. T

agentsdair-ai--x
1 May 2026
Industry

A working paper uses an LLM to analyze political discourse on X, finding that anger is the dominant emotion expressed by US users, especially those over 65 (Tim Harford/Financial Times)

DGX agent

Tim Harford / Financial Times: A working paper uses an LLM to analyze political discourse on X, finding that anger is the dominant emotion expressed by US users, especially those over 65 — There is a

industrytechmeme
29 Apr 2026
Tutorials

Ambient diffusion was a really cool paper in this space. Given corrupted data, corrupt even further, such that the model can’t infer what wa…

DGX agent

Ambient diffusion was a really cool paper in this space. Given corrupted data, corrupt even further, such that the model can’t infer what was genuine corruption vs synthetically added. In expectation

tutorialslinus-lee--x
29 Apr 2026
Tools

eugene is doing a great job covering it on the @latentspacepod paper club (will post on the youtube soon) https://x.com/picocreator/status/2…

DGX agent

eugene is doing a great job covering it on the @latentspacepod paper club (will post on the youtube soon) https://x.com/picocreator/status/2047626012419313889?s=20 🗒️ Other things - FP4 MoE training -

toolsswyx--x
29 Apr 2026
Industry

Research paper: https://www.pnas.org/doi/abs/10.1073/pnas.2531743123#abstract Original video : https://www.youtube.com/watch?v=jli0jPiKxQs I…

DGX agent

Research paper: https://www.pnas.org/doi/abs/10.1073/pnas.2531743123#abstract Original video : https://www.youtube.com/watch?v=jli0jPiKxQs I break down stories like this every day in my free newslette

industryrowan-cheung--x
29 Apr 2026
Model Releases

The new DeepSeek-V4, like DeepSeek-V3, uses concepts from our 2024 paper on Self-Rewarding LMs -- see screenshots of their tech reports! (Co…

DGX agent

The new DeepSeek-V4, like DeepSeek-V3, uses concepts from our 2024 paper on Self-Rewarding LMs -- see screenshots of their tech reports! (Congrats!) Classical :) Self-Rewarding LMs from Jan 2024: http

model-releasesjeremy-howard--x
29 Apr 2026
Agents

.@huggingface unveiled ml-intern – an open-source agent that automates the gritty post-training loop: - reading papers - tracing citations -…

DGX agent

.@huggingface unveiled ml-intern – an open-source agent that automates the gritty post-training loop: - reading papers - tracing citations - curating datasets - running experiments - and iterating lik

agentsclem-delangue--x
28 Apr 2026
Applications

The problem is that the incentives push for 'more' over 'better' Paper: https://pubsonline.informs.org/doi/full/10.1287/orsc.2026.ed.v37.n3

DGX agent

This paper examines how organizational incentive structures often prioritize quantitative growth and output volume over quality improvements, creating misaligned motivations that can undermine long-te

applicationsethan-mollick--x
27 Apr 2026
Model Releases

The Top AI Papers of the Week (April 19 - 26) - Skill-RAG - DeepSeek V4 - Autogenesis - Attention to Mamba - Stateless Decision Memory - Sel…

DGX agent

The Top AI Papers of the Week (April 19 - 26) - Skill-RAG - DeepSeek V4 - Autogenesis - Attention to Mamba - Stateless Decision Memory - Self-Evolving Logic Synthesis - Self-Generated World Knowledge

model-releasesdair-ai--x
26 Apr 2026
Applications

There is also too much focus on two issues in discussions of paper reviewing: hallucinations & privacy. Hallucinations are not gone, but the…

DGX agent

There is also too much focus on two issues in discussions of paper reviewing: hallucinations & privacy. Hallucinations are not gone, but the latest models rarely hallucinate sources (& it is relativel

applicationsethan-mollick--x
25 Apr 2026
Agents

DiagramBank: A Large-scale Dataset of Diagram Design Exemplars with Paper Metadata for Retrieval-Augmented Generation

DGX agent

arXiv:2604.20857v1 Announce Type: cross Abstract: Recent advances in autonomous ``AI scientist'' systems have demonstrated the ability to automatically write scientific manuscripts and codes with exec

agentsarxiv-cs-ai
24 Apr 2026
Safety

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

DGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

safetyarxiv-cs-cl
24 Apr 2026
Model Releases

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds resea…

DGX agent

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds research-backed models autonomously. Pushed a benchmark from 10%

model-releasesclem-delangue--x
21 Apr 2026
Agents

I think the CoALA paper's classification system of semantic/episodic/procedural is maybe the closest thing we have to a standard for agent m…

DGX agent

The CoALA paper proposes a classification system for agent memory that distinguishes between semantic memory (facts and concepts), episodic memory (specific experiences and events), and procedural mem

agentsharrison-chase--x
21 Apr 2026
Model Releases

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Co…

DGX agent

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Codex land near the human median, but with far tighter dispers

model-releasesethan-mollick--x
20 Apr 2026
Industry

First paper, microsite & NotebookLM https://ii.inc/web/releases/one-postulate

DGX agent

Emad Mostaque announced the release of Stability AI's first paper, microsite, and integration with NotebookLM, likely detailing new research findings or technical capabilities. The announcement appear

industryemad-mostaque--x
20 Apr 2026
Model Releases

The Top AI Papers of the Week (April 13 - 19) - AlphaEval - AiScientist - Auto-Diagnose - Nemotron 3 Super - Subliminal Learning - Automated…

DGX agent

The Top AI Papers of the Week (April 13 - 19) - AlphaEval - AiScientist - Auto-Diagnose - Nemotron 3 Super - Subliminal Learning - Automated W2S Researcher - Memory Transfer Learning Read on for more:

model-releasesdair-ai--x
19 Apr 2026
Model Releases

A lot of papers coming out are still focused on GPT-4, but you could extrapolate their effects to GPT-5, etc. Much harder to know what the i…

DGX agent

A lot of papers coming out are still focused on GPT-4, but you could extrapolate their effects to GPT-5, etc. Much harder to know what the impacts of Claude Code/Codex etc. are because they are so new

model-releasesethan-mollick--x
16 Apr 2026
Research

Was looking at a ICLR 2025 Oral paper and I am shocked it got oral [D]

DGX agent

This Reddit thread from r/MachineLearning reflects community skepticism about the peer review standards at top ML conferences, with a user expressing surprise that a particular paper received an oral

researchr-machinelearning
15 Apr 2026
Model Releases

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

DGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

model-releasesarxiv-cs-ai
14 Apr 2026
Industry

Research paper link: https://www.pnas.org/doi/10.1073/pnas.2519129123 I break down stories like this every day in my free newsletter. Keep u…

DGX agent

Research paper link: https://www.pnas.org/doi/10.1073/pnas.2519129123 I break down stories like this every day in my free newsletter. Keep up with the latest in AI/Robotics in 5 min a day: https://www

industryrowan-cheung--x
14 Apr 2026
Industry

Research paper link: https://www.cell.com/cell/fulltext/S0092-8674%2825%2901312-1 I break down stories like this every day in my free newsle…

DGX agent

Research paper link: https://www.cell.com/cell/fulltext/S0092-8674%2825%2901312-1 I break down stories like this every day in my free newsletter. Keep up with the latest in AI/Robotics in 5 min a day:

industryrowan-cheung--x
13 Apr 2026
Agents

The Top AI Papers of the Week (April 6 - 12) - Memento - Neural Computers - The Universal Verifier - Agent Skills in the Wild - Memory Intel…

DGX agent

The Top AI Papers of the Week (April 6 - 12) - Memento - Neural Computers - The Universal Verifier - Agent Skills in the Wild - Memory Intelligence Agent (MIA) - Single-Agent vs Multi-Agent LLMs - Sca

agentsdair-ai--x
12 Apr 2026
Model Releases

I’ve uploaded a new paper on arXiv (co-authored by @rasbt): MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning In Parameter-Efficient…

DGX agent

I’ve uploaded a new paper on arXiv (co-authored by @rasbt): MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning In Parameter-Efficient Fine-Tuning, a key question may not just be how low-rank th

model-releasessebastian-raschka--x
8 Apr 2026
Applications

Some approaches: 1) Multiple reviews. Some papers already show having many AI team members review a problem reduces errors 2) Building in te…

DGX agent

Some approaches: 1) Multiple reviews. Some papers already show having many AI team members review a problem reduces errors 2) Building in tests and checkpoints 3) Multiple independent answers that are

applicationsethan-mollick--x
8 Apr 2026
Model Releases

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimo…

DGX agent

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimodal generative modeling could accelerate protein sciences! ⬇

model-releasesyoshua-bengio--x
8 Apr 2026
Research

Top Community Contributors: @SHL0MS (7 PRs) — p5js creative coding skill, manim-video skill + 5 reference expansions, research-paper-writing…

DGX agent

Top Community Contributors: @SHL0MS (7 PRs) — p5js creative coding skill, manim-video skill + 5 reference expansions, research-paper-writing, Nous OAuth fix, manim fix @sidbing (3 PRs) — Firecrawl clo

researchnous-research--x
8 Apr 2026
Applications

More on the @nytimes piece about that $1.8B, two-person, AI company ... not the paper's finest moment in quick retrospect. And in the health…

DGX agent

More on the @nytimes piece about that $1.8B, two-person, AI company ... not the paper's finest moment in quick retrospect. And in the healthcare context no less. Our friend @GaryMarcus was on this a f

applicationsgary-marcus--x
7 Apr 2026
Model Releases

[PAPER] GPQA, MMLU-Pro, and MMMU-Pro were audited for broken questions, and up to 12% of them had to be removed. New drop in clean versions released

DGX agent

I was very curious why all the models were topping out on GPQA-Diamond around 92 or 93% (AA) and spent the last few weeks pouring over GPQA (Diamond and Extended), and then expanded to auditing MMLU-P

model-releasesr-localllama
28 Jul 2026
Model Releases

Towards Automating Scientific Review with Google's Paper Assistant Tool

DGX agent

arXiv:2606.28277v1 Announce Type: cross Abstract: Artificial intelligence is driving a revolution in scientific discovery, accelerating everything from hypothesis generation to mathematical theorem pr

model-releasesarxiv-cs-ai
29 Jun 2026
Model Releases

In policy paper, OpenAI diverges from White House on AI safety

DGX agent

OpenAI Group PBC’s newly released proposal for how advanced artificial intelligence should be regulated differs slightly from the Trump administration’s executive order, also released this week. Relea

model-releasessiliconangle
4 Jun 2026
Model Releases

HakushoBench: A Japanese Chart and Table VQA Benchmark from Governmental White Papers

DGX agent

arXiv:2606.01132v1 Announce Type: new Abstract: Understanding chart and table images is essential for applying vision-language models (VLMs) to real-world document understanding. While English benchma

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers

DGX agent

arXiv:2605.25415v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in academic peer review, yet their reliability, alignment with human judgment, and robustness to adve

model-releasesarxiv-cs-cl
26 May 2026
Applications

FlowPlan-G2P: A Structured Generation Framework for Transforming Scientific Papers into Patent Descriptions

DGX agent

arXiv:2601.02589v2 Announce Type: replace-cross Abstract: Over 3.5 million patents are filed annually, with drafting patent descriptions requiring deep technical and legal expertise. Transforming scie

applicationsarxiv-cs-ai
15 Apr 2026
Model Releases

GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses

DGX agent

arXiv:2604.11924v1 Announce Type: new Abstract: While LLMs hold significant potential to transform scientific research, we advocate for their use to augment and empower researchers rather than to auto

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

NovBench: Evaluating Large Language Models on Academic Paper Novelty Assessment

DGX agent

arXiv:2604.11543v1 Announce Type: cross Abstract: Novelty is a core requirement in academic publishing and a central focus of peer review, yet the growing volume of submissions has placed increasing p

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

PosterGen: Aesthetic-Aware Multi-Modal Paper-to-Poster Generation via Multi-Agent LLMs

DGX agent

arXiv:2508.17188v2 Announce Type: replace Abstract: Multi-agent systems built upon large language models (LLMs) have demonstrated remarkable capabilities in tackling complex compositional tasks. In th

agentsarxiv-cs-ai
14 Apr 2026
Agents

Working Paper: Towards Schema-based Learning from a Category-Theoretic Perspective

DGX agent

arXiv:2604.10589v1 Announce Type: new Abstract: We introduce a hierarchical categorical framework for Schema-Based Learning (SBL) structured across four interconnected levels. At the schema level, a f

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation

DGX agent

arXiv:2504.10284v5 Announce Type: replace Abstract: Literature review tables are essential for summarizing and comparing collections of scientific papers. In this paper, we study the automatic generat

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

Working Paper: Towards a Category-theoretic Comparative Framework for Artificial General Intelligence

DGX agent

arXiv:2603.28906v2 Announce Type: replace Abstract: AGI has become the Holly Grail of AI with the promise of level intelligence and the major Tech companies around the world are investing unprecedente

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

How many people in this sub try to train their own AI from scratch on their systems just for fun and to test out techniques from research papers?

DGX agent

As for me, I own a system with an RTX 5090, Ryzen 9 9950X3D2, and 64 GB of DDR5. Every time I see research come out with a new way to train AI, I immediately think to try it on my system to see the re

model-releasesr-localllama
6 Aug 2026
Model Releases

LiveXiv -- A Multi-Modal Live Benchmark Based on Arxiv Papers Content

DGX agent

arXiv:2410.10783v4 Announce Type: replace Abstract: The large-scale training of multi-modal models on data scraped from the web has shown outstanding utility in infusing these models with the required

model-releasesarxiv-cs-cv
6 Aug 2026
← Previous
1…56789…252
Next →