AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,056 results
28 Apr 2026

.@huggingface unveiled ml-intern – an open-source agent that automates the gritty post-training loop: - reading papers - tracing citations -…

AgentsDGX agent

.@huggingface unveiled ml-intern – an open-source agent that automates the gritty post-training loop: - reading papers - tracing citations - curating datasets - running experiments - and iterating lik

The Last Human-Written Paper: Agent-Native Research Artifacts

AgentsDGX agent

arXiv:2604.24658v1 Announce Type: new Abstract: Scientific publication compresses a branching, iterative research process into a linear narrative, discarding the majority of what was discovered along

27 Apr 2026

The problem is that the incentives push for 'more' over 'better' Paper: https://pubsonline.informs.org/doi/full/10.1287/orsc.2026.ed.v37.n3

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
ApplicationsDGX agent

This paper examines how organizational incentive structures often prioritize quantitative growth and output volume over quality improvements, creating misaligned motivations that can undermine long-te

// Agentic World Modeling // Massive 40-author survey just dropped. Cleanest taxonomy of world models in agent research I've seen. (bookmark…

AgentsDGX agent

// Agentic World Modeling // Massive 40-author survey just dropped. Cleanest taxonomy of world models in agent research I've seen. (bookmark it) The paper proposes a 'levels × laws' framework. Three c

26 Apr 2026

The Top AI Papers of the Week (April 19 - 26) - Skill-RAG - DeepSeek V4 - Autogenesis - Attention to Mamba - Stateless Decision Memory - Sel…

Model ReleasesDGX agent

The Top AI Papers of the Week (April 19 - 26) - Skill-RAG - DeepSeek V4 - Autogenesis - Attention to Mamba - Stateless Decision Memory - Self-Evolving Logic Synthesis - Self-Generated World Knowledge

25 Apr 2026

There is also too much focus on two issues in discussions of paper reviewing: hallucinations & privacy. Hallucinations are not gone, but the…

ApplicationsDGX agent

There is also too much focus on two issues in discussions of paper reviewing: hallucinations & privacy. Hallucinations are not gone, but the latest models rarely hallucinate sources (& it is relativel

24 Apr 2026

DiagramBank: A Large-scale Dataset of Diagram Design Exemplars with Paper Metadata for Retrieval-Augmented Generation

AgentsDGX agent

arXiv:2604.20857v1 Announce Type: cross Abstract: Recent advances in autonomous ``AI scientist'' systems have demonstrated the ability to automatically write scientific manuscripts and codes with exec

XtraGPT: Context-Aware and Controllable Academic Paper Revision via Human-AI Collaboration

SafetyDGX agent

arXiv:2505.11336v4 Announce Type: replace Abstract: Despite the growing adoption of large language models (LLMs) in academic workflows, their capabilities remain limited in supporting high-quality sci

21 Apr 2026

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds resea…

Model ReleasesDGX agent

HUGGING FACE JUST AUTOMATED THEIR ENTIRE POST-TRAINING TEAM WITH AN AGENT. It reads papers, runs GPU experiments, iterates, and builds research-backed models autonomously. Pushed a benchmark from 10%

I think the CoALA paper's classification system of semantic/episodic/procedural is maybe the closest thing we have to a standard for agent m…

AgentsDGX agent

The CoALA paper proposes a classification system for agent memory that distinguishes between semantic memory (facts and concepts), episodic memory (specific experiences and events), and procedural mem

20 Apr 2026

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Co…

Model ReleasesDGX agent

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Codex land near the human median, but with far tighter dispers

First paper, microsite & NotebookLM https://ii.inc/web/releases/one-postulate

IndustryDGX agent

Emad Mostaque announced the release of Stability AI's first paper, microsite, and integration with NotebookLM, likely detailing new research findings or technical capabilities. The announcement appear

19 Apr 2026

The Top AI Papers of the Week (April 13 - 19) - AlphaEval - AiScientist - Auto-Diagnose - Nemotron 3 Super - Subliminal Learning - Automated…

Model ReleasesDGX agent

The Top AI Papers of the Week (April 13 - 19) - AlphaEval - AiScientist - Auto-Diagnose - Nemotron 3 Super - Subliminal Learning - Automated W2S Researcher - Memory Transfer Learning Read on for more:

16 Apr 2026

A lot of papers coming out are still focused on GPT-4, but you could extrapolate their effects to GPT-5, etc. Much harder to know what the i…

Model ReleasesDGX agent

A lot of papers coming out are still focused on GPT-4, but you could extrapolate their effects to GPT-5, etc. Much harder to know what the impacts of Claude Code/Codex etc. are because they are so new

15 Apr 2026

Was looking at a ICLR 2025 Oral paper and I am shocked it got oral [D]

ResearchDGX agent

This Reddit thread from r/MachineLearning reflects community skepticism about the peer review standards at top ML conferences, with a user expressing surprise that a particular paper received an oral

FlowPlan-G2P: A Structured Generation Framework for Transforming Scientific Papers into Patent Descriptions

ApplicationsDGX agent

arXiv:2601.02589v2 Announce Type: replace-cross Abstract: Over 3.5 million patents are filed annually, with drafting patent descriptions requiring deep technical and legal expertise. Transforming scie

GoodPoint: Learning Constructive Scientific Paper Feedback from Author Responses

Model ReleasesDGX agent

arXiv:2604.11924v1 Announce Type: new Abstract: While LLMs hold significant potential to transform scientific research, we advocate for their use to augment and empower researchers rather than to auto

14 Apr 2026

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

Model ReleasesDGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

Research paper link: https://www.pnas.org/doi/10.1073/pnas.2519129123 I break down stories like this every day in my free newsletter. Keep u…

IndustryDGX agent

Research paper link: https://www.pnas.org/doi/10.1073/pnas.2519129123 I break down stories like this every day in my free newsletter. Keep up with the latest in AI/Robotics in 5 min a day: https://www

NovBench: Evaluating Large Language Models on Academic Paper Novelty Assessment

Model ReleasesDGX agent

arXiv:2604.11543v1 Announce Type: cross Abstract: Novelty is a core requirement in academic publishing and a central focus of peer review, yet the growing volume of submissions has placed increasing p

PosterGen: Aesthetic-Aware Multi-Modal Paper-to-Poster Generation via Multi-Agent LLMs

AgentsDGX agent

arXiv:2508.17188v2 Announce Type: replace Abstract: Multi-agent systems built upon large language models (LLMs) have demonstrated remarkable capabilities in tackling complex compositional tasks. In th

Working Paper: Towards Schema-based Learning from a Category-Theoretic Perspective

AgentsDGX agent

arXiv:2604.10589v1 Announce Type: new Abstract: We introduce a hierarchical categorical framework for Schema-Based Learning (SBL) structured across four interconnected levels. At the schema level, a f

13 Apr 2026

Research paper link: https://www.cell.com/cell/fulltext/S0092-8674%2825%2901312-1 I break down stories like this every day in my free newsle…

IndustryDGX agent

Research paper link: https://www.cell.com/cell/fulltext/S0092-8674%2825%2901312-1 I break down stories like this every day in my free newsletter. Keep up with the latest in AI/Robotics in 5 min a day:

12 Apr 2026

The Top AI Papers of the Week (April 6 - 12) - Memento - Neural Computers - The Universal Verifier - Agent Skills in the Wild - Memory Intel…

AgentsDGX agent

The Top AI Papers of the Week (April 6 - 12) - Memento - Neural Computers - The Universal Verifier - Agent Skills in the Wild - Memory Intelligence Agent (MIA) - Single-Agent vs Multi-Agent LLMs - Sca

8 Apr 2026

I’ve uploaded a new paper on arXiv (co-authored by @rasbt): MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning In Parameter-Efficient…

Model ReleasesDGX agent

I’ve uploaded a new paper on arXiv (co-authored by @rasbt): MiCA Learns More Knowledge Than LoRA and Full Fine-Tuning In Parameter-Efficient Fine-Tuning, a key question may not just be how low-rank th

Some approaches: 1) Multiple reviews. Some papers already show having many AI team members review a problem reduces errors 2) Building in te…

ApplicationsDGX agent

Some approaches: 1) Multiple reviews. Some papers already show having many AI team members review a problem reduces errors 2) Building in tests and checkpoints 3) Multiple independent answers that are

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimo…

Model ReleasesDGX agent

Today, our group at @Mila_Quebec and the lab of @francesarnold at @Caltech just released a new paper I contributed to, exploring how multimodal generative modeling could accelerate protein sciences! ⬇

Top Community Contributors: @SHL0MS (7 PRs) — p5js creative coding skill, manim-video skill + 5 reference expansions, research-paper-writing…

ResearchDGX agent

Top Community Contributors: @SHL0MS (7 PRs) — p5js creative coding skill, manim-video skill + 5 reference expansions, research-paper-writing, Nous OAuth fix, manim fix @sidbing (3 PRs) — Firecrawl clo

7 Apr 2026

More on the @nytimes piece about that $1.8B, two-person, AI company ... not the paper's finest moment in quick retrospect. And in the health…

ApplicationsDGX agent

More on the @nytimes piece about that $1.8B, two-person, AI company ... not the paper's finest moment in quick retrospect. And in the healthcare context no less. Our friend @GaryMarcus was on this a f

28 Jul 2026

[PAPER] GPQA, MMLU-Pro, and MMMU-Pro were audited for broken questions, and up to 12% of them had to be removed. New drop in clean versions released

Model ReleasesDGX agent

I was very curious why all the models were topping out on GPQA-Diamond around 92 or 93% (AA) and spent the last few weeks pouring over GPQA (Diamond and Extended), and then expanded to auditing MMLU-P

PNAS: Over Half of All Academic Articles Now Show LLM Influence—7.3M-Paper Study [R]

SafetyDGX agent

Largest empirical study of AI penetration in academic publishing ever conducted—51%-by-2025 is the most authoritative quantitative marker yet of how thoroughly LLMs have reshaped scientific writing, a

29 Jun 2026

Towards Automating Scientific Review with Google's Paper Assistant Tool

Model ReleasesDGX agent

arXiv:2606.28277v1 Announce Type: cross Abstract: Artificial intelligence is driving a revolution in scientific discovery, accelerating everything from hypothesis generation to mathematical theorem pr

4 Jun 2026

In policy paper, OpenAI diverges from White House on AI safety

Model ReleasesDGX agent

OpenAI Group PBC’s newly released proposal for how advanced artificial intelligence should be regulated differs slightly from the Trump administration’s executive order, also released this week. Relea

2 Jun 2026

HakushoBench: A Japanese Chart and Table VQA Benchmark from Governmental White Papers

Model ReleasesDGX agent

arXiv:2606.01132v1 Announce Type: new Abstract: Understanding chart and table images is essential for applying vision-language models (VLMs) to real-world document understanding. While English benchma

26 May 2026

LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers

Model ReleasesDGX agent

arXiv:2605.25415v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in academic peer review, yet their reliability, alignment with human judgment, and robustness to adve

10 Apr 2026

arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation

Model ReleasesDGX agent

arXiv:2504.10284v5 Announce Type: replace Abstract: Literature review tables are essential for summarizing and comparing collections of scientific papers. In this paper, we study the automatic generat

Working Paper: Towards a Category-theoretic Comparative Framework for Artificial General Intelligence

AgentsDGX agent

arXiv:2603.28906v2 Announce Type: replace Abstract: AGI has become the Holly Grail of AI with the promise of level intelligence and the major Tech companies around the world are investing unprecedente

6 Aug 2026

How many people in this sub try to train their own AI from scratch on their systems just for fun and to test out techniques from research papers?

Model ReleasesDGX agent

As for me, I own a system with an RTX 5090, Ryzen 9 9950X3D2, and 64 GB of DDR5. Every time I see research come out with a new way to train AI, I immediately think to try it on my system to see the re

LiveXiv -- A Multi-Modal Live Benchmark Based on Arxiv Papers Content

Model ReleasesDGX agent

arXiv:2410.10783v4 Announce Type: replace Abstract: The large-scale training of multi-modal models on data scraped from the web has shown outstanding utility in infusing these models with the required

5 Aug 2026

A machine-readable catalogue of the Tsiolkovsky papers (fond 555, Archive of the Russian Academy of Sciences), and a way to measure how well its handwriting can be read

ResearchDGX agent

arXiv:2608.03617v1 Announce Type: new Abstract: The personal archive of Konstantin Tsiolkovsky (1857-1935) is held as fond 555 of the Archive of the Russian Academy of Sciences. The archive scanned th

26 Jul 2026

[Paper] RecGPT-V3 Technical Report

Local AiDGX agent

Large language models (LLMs) are transforming recommender systems from matching co-occurrence patterns in historical behavior toward reasoning about the intent that drives it. RecGPT-V1 pioneered this

23 Jul 2026

[Paper] SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

Model ReleasesDGX agent

Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed training, including severe memory pressure, non-overlappe

1 Jul 2026

4/ Escaping the Verifier: Learning to Reason via Demonstrations (RARO) Paper: https://arxiv.org/abs/2511.21667

ToolsDGX agent

RARO (Reasoning via Demonstrations) is a method for training AI models to improve reasoning capabilities by learning from demonstrations rather than relying solely on external verifiers. The approach

1/ DSGym: A Holistic Framework for Evaluating and Training Data Science Agents Paper: https://arxiv.org/abs/2601.16344

ToolsDGX agent

DSGym is a comprehensive framework designed to evaluate and train AI agents for data science tasks, providing a structured environment for benchmarking agent performance across various data science wo

2/ ThunderAgent: A Simple, Fast and Program-Aware Agentic Inference System Paper: https://arxiv.org/abs/2602.13692

AgentsDGX agent

ThunderAgent is an agentic inference system designed for fast and efficient execution of AI agent programs, developed by Together AI. The system appears to optimize program-aware inference by leveragi

30 Jun 2026

Summary: TGT’s 2026 ICML Papers

SafetyDGX agent

The International Conference on Machine Learning (ICML), held annually for over forty years, is among the most influential conferences in modern AI research. This year in Seoul, ICML is hosting its se

Supporting Workflow Reproducibility by Linking Bioinformatics Tools across Papers and Executable Code

ResearchDGX agent

arXiv:2603.08195v2 Announce Type: replace Abstract: Motivation: The rapid growth of biological data has intensified the need for transparent, reproducible, and well-documented computational workflows.

Temporal Posed and Spontaneous Gesture Recognition from Electromyography in the Rock-Paper-Scissors Game

ApplicationsDGX agent

arXiv:2606.29423v1 Announce Type: new Abstract: The importance of gesture recognition has been acknowledged in many domains requiring real-time recognition systems. Two requirements for these are fast

26 Jun 2026

SciFig: Towards Automating Editable Figure Generation for Scientific Papers

Model ReleasesDGX agent

arXiv:2601.04390v2 Announce Type: replace Abstract: High-quality methodology figures are central to scientific communication, yet they remain difficult and time-consuming to create. Such figures must

24 Jun 2026

NatureBench: Can Coding Agents Match the Published SOTA of Nature-Family Papers?

Model ReleasesDGX agent

arXiv:2606.24530v1 Announce Type: new Abstract: We introduce NatureBench, a cross-discipline benchmark of 90 tasks distilled from peer-reviewed Nature-family publications, designed to evaluate whether

8 Jun 2026

ChemQuests: A Curated Chemistry Question-Answer Database Extracted from ChemRxiv papers

ApplicationsDGX agent

arXiv:2505.05232v3 Announce Type: replace Abstract: The rapid expansion of chemistry literature poses significant challenges for researchers seeking to efficiently access domain-specific knowledge. To

28 May 2026

Are We Truly Innovating? A Qualitative and Quantitative Study of Originality in AI Research Papers

ResearchDGX agent

arXiv:2602.06054v3 Announce Type: replace Abstract: Assessing originality in AI research is arguably the most consequential yet least reliable step in peer review. Reviewer judgments of originality re

23 May 2026

receipts for most points can found here, if you read this paper closely: https://nautil.us/deep-learning-is-hitting-a-wall-238440

SafetyDGX agent

Gary Marcus references a Nautilus article arguing that deep learning is encountering fundamental limitations, suggesting readers can find supporting evidence and detailed arguments for this perspectiv

15 May 2026

Paper: https://arxiv.org/abs/2605.06554 Code: https://github.com/ighoshsubho/lighthouse-attention HF: https://huggingface.co/papers/2605.065…

ResearchDGX agent

Lighthouse Attention is a novel attention mechanism that improves efficiency in transformer models by selectively focusing computation on the most relevant tokens, similar to how a lighthouse beam ill

11 May 2026

Self Driving Datasets: From 20 Million Papers to Nuanced Biomedical Knowledge at Scale

Local AiDGX agent

arXiv:2605.07022v1 Announce Type: new Abstract: Manually curated biomedical repositories -- spanning bioactivity, genomics, and chemistry -- are expensive to maintain, lag behind primary literature, a

23 Apr 2026

ICLR 2026: 12 papers on making AI systems reliable, efficient, and secure

SafetyDGX agent

A 7B agent that beats GPT-4o. Lossless weight compression that speeds up inference by 177%. An arena where 23 teams battled across 103,000 adversarial rounds. This year at ICLR, Lambda is presenting t

12 Aug 2026

Eleven Years of BRACIS: A Meta-Scientific Study of the Brazilian Conference on Intelligent Systems

ResearchDGX agent

arXiv:2608.09964v1 Announce Type: cross Abstract: The Brazilian Conference on Intelligent Systems (BRACIS) is the main national venue for Artificial Intelligence research in Brazil, hosted by the Braz

10 Aug 2026

LitTraceQA: A Benchmark for Multi-Stage Grounding and Verification in Scientific Question Answering

Model ReleasesDGX agent

arXiv:2608.07370v1 Announce Type: new Abstract: Scientific literature is increasingly used as a knowledge source for language models, retrieval-augmented generation systems, and research assistants, b

7 Aug 2026

Hijacking Robots with a Piece of Paper: A Systematic Study of Physical Prompt Injection in VLM-Controlled Robots

Model ReleasesDGX agent

arXiv:2608.05715v1 Announce Type: cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as planners in robotic systems, where they translate natural-language commands into executable

19 Jul 2026

Wiki Lint Report — 2026-07-19

SynthesesDGX agent

Automated lint: 20 errors, 8743 warnings, 3 info

← Previous
1…45678…201
Next →