AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,056 results
1 Jul 2026

3/ Learning to Discover at Test Time (TTT-Discover) Paper: https://arxiv.org/abs/2601.16175

ToolsDGX agent

TTT-Discover is a method that enables models to learn and discover patterns during test time rather than only during training, allowing for adaptation to new data distributions at inference. The appro

6/ When RL Meets Adaptive Speculative Training: A Unified Training-Serving System (Aurora) Paper: https://arxiv.org/abs/2602.06932

ToolsDGX agent

Aurora is a unified training-serving system that integrates reinforcement learning with adaptive speculative training to optimize large language model inference and training efficiency. The system dyn

7/ Untied Ulysses: Memory-Efficient Context Parallelism via Headwise Chunking Paper: https://arxiv.org/abs/2602.21196

ToolsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Untied Ulysses is a memory-efficient technique for context parallelism that processes attention heads in chunks rather than sequences, reducing memory overhead during transformer inference and trainin

8/ Opportunistic Expert Activation: Batch-Aware Expert Routing for Faster Decode Without Retraining (OEA) Paper: https://arxiv.org/abs/2511.…

ToolsDGX agent

Opportunistic Expert Activation (OEA) is a batch-aware expert routing technique for mixture-of-experts models that enables faster decoding without requiring model retraining. The method optimizes whic

9/ ParallelKernelBench: Benchmarking LLMs on Multi-GPU Kernel Generation Paper: https://www.alphaxiv.org/abs/2606.parallel-kernel-bench

HardwareDGX agent

ParallelKernelBench is a benchmarking framework designed to evaluate large language models' ability to generate optimized GPU kernels for multi-GPU computing environments. The benchmark assesses LLMs

23 Jun 2026

Read the technical paper on Krea 2 https://www.krea.ai/blog/krea-2-technical-report Download the model weights https://github.com/krea-ai/kr…

Local AiDGX agent

Krea 2 is a technical advancement in AI image generation with newly released model weights available for download on GitHub. The technical report details the improvements and capabilities of this vers

2 Jun 2026

Position Paper: Post-Solve Robustness in Decision Engines: Feasible Regions and Smoothness Under Perturbations

Model ReleasesDGX agent

arXiv:2606.00002v1 Announce Type: new Abstract: Mixed-Integer Linear Programming (MILP) decision engines routinely output nominally optimal plans for high-stakes industrial systems. Yet deployment rar

I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications

Model ReleasesDGX agent

arXiv:2606.00750v1 Announce Type: new Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, existing

OARelatedWork: A Large-Scale Dataset of Related Work Sections with Full-texts from Open Access Sources

Model ReleasesDGX agent

arXiv:2405.01930v2 Announce Type: replace Abstract: This paper introduces OARelatedWork: a dataset for related work generation from open-access sources. It is the first large-scale multi-document summ

PaperVoyager : Building Interactive Web with Visual Language Models

Model ReleasesDGX agent

arXiv:2603.22999v3 Announce Type: replace Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, exist

Who Annotates in NLP? A Large-scale Assessment of Human Annotation Reporting between 2018 and 2025

ResearchDGX agent

arXiv:2606.02255v1 Announce Type: cross Abstract: Human annotation is the empirical foundation of much NLP research, from dataset construction to model evaluation, but papers often leave unclear who p

26 May 2026

a shout out to the paper: https://arxiv.org/html/2605.25376v1 'KYA: A Framework-Agnostic Trust Layer for Autonomous Systems with Verifiable …

SafetyDGX agent

KYA is a framework-agnostic trust layer designed for autonomous systems that provides verifiable guarantees, addressing the need for trustworthy and transparent operation of AI agents across different

LECTOR: Joint Optimization of Scientific Reasoning Graphs and Introduction Generation

ResearchDGX agent

arXiv:2605.25964v1 Announce Type: new Abstract: AI Scientists have shown promising progress across multiple stages of the research pipeline, among which automatic scientific paper writing remains a fo

Methods for Formal Verification of Agent Skills: Three Layers Toward a Mechanically Checkable Capability-Containment Proof

AgentsDGX agent

arXiv:2605.23951v1 Announce Type: new Abstract: The companion paper introduced a four-level verification lattice on agent-skill manifests (unverified, declared, tested, formal) and left the top level

14 May 2026

VERA-MH Concept Paper

Model ReleasesDGX agent

arXiv:2510.15297v4 Announce Type: replace-cross Abstract: We introduce VERA-MH (Validation of Ethical and Responsible AI in Mental Health), an automated evaluation of the safety of AI chatbots used in

6 May 2026

SCION: Size-aware Policy Orchestration for Nonstationary Object Caches (Long Paper Version)

SafetyDGX agent

arXiv:2605.01055v1 Announce Type: cross Abstract: Object caches underpin cloud and edge services, but production workloads are heterogeneous, nonstationary, and throughput-constrained. Recent simple n

28 Apr 2026

Learning to Orchestrate Agents in Natural Language with the Conductor Fugu Blog: https://sakana.ai/fugu-beta Paper: https://arxiv.org/abs/25…

ResearchDGX agent

Conductor is a method for orchestrating multiple AI agents through natural language instructions, enabling coordinated multi-agent systems where a central 'conductor' agent directs specialized agents

21 Apr 2026

HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution

Model ReleasesDGX agent

arXiv:2604.17745v1 Announce Type: new Abstract: Recent advances in large language models have highlighted their potential to automate computational research, particularly reproducing experimental resu

18 Apr 2026

Paper from Kimi: Prefill-as-a-Service: KVCache of Next-Generation Models Could Go Cross-Datacenter [R]

ResearchDGX agent

Mooncake is the serving platform for Kimi developed by Moonshot AI, featuring a KVCache-centric disaggregated architecture that separates prefill and decoding clusters while leveraging underutilized C

13 Apr 2026

From Paper to Program: Accelerating Quantum Many-Body Algorithm Development via a Multi-Stage LLM-Assisted Workflow

Model ReleasesDGX agent

arXiv:2604.04089v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can generate code rapidly but remain unreliable for scientific algorithms whose correctness depends on structural

Mandatory In-Person Presentation in CVPR 2026 [D]

ResearchDGX agent

This Reddit thread on r/MachineLearning discusses CVPR 2026's policy requiring accepted papers to be registered under an in-person author registration, with virtual attendance still permitted if circu

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences

Model ReleasesDGX agent

arXiv:2602.11354v2 Announce Type: replace Abstract: The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on

10 Apr 2026

New paper argues history, not mantle plume, powers Yellowstone

IndustryDGX agent

A study published in *Science* (April 2026) by researchers from the Chinese Academy of Sciences challenges the long-held view that Yellowstone's supervolcano is powered by a deep mantle plume, inst...

30 Jun 2026

Unveiling Novelty Evolution in the field of Library and Information Science in China

ResearchDGX agent

arXiv:2606.29872v1 Announce Type: cross Abstract: This study analyzes the novelty distribution of scholarly papers in the field of Library and Information Science (LIS) in China, with a focus on diffe

29 May 2026

ChatGPT diagnosed 40 million people with a disease that was invented as a joke. Not a real disease. Not a misunderstood disease. A completel…

Model ReleasesDGX agent

ChatGPT diagnosed 40 million people with a disease that was invented as a joke. Not a real disease. Not a misunderstood disease. A completely fictional condition with a fake name, fake papers, and fak

27 May 2026

ANNOUNCEMENT: WE’RE SAVING SCIENCE! We’re often told that science is “self-correcting.” But that’s not really true. Science doesn’t correct …

IndustryDGX agent

ANNOUNCEMENT: WE’RE SAVING SCIENCE! We’re often told that science is “self-correcting.” But that’s not really true. Science doesn’t correct itself like a thermostat adjusting the temperature in your h

20 May 2026

How Far Are We From True Auto-Research?

Model ReleasesDGX agent

arXiv:2605.19156v1 Announce Type: new Abstract: Recent auto-research systems can produce complete papers, but feasibility is not the same as quality, and the field still lacks a systematic study of ho

12 May 2026

Can Deep Research Agents Retrieve and Organize? Evaluating the Synthesis Gap with Expert Taxonomies

Model ReleasesDGX agent

arXiv:2601.12369v3 Announce Type: replace Abstract: Deep Research Agents increasingly automate survey generation, yet whether they match human experts at retrieving essential papers and organizing the

Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities

Model ReleasesDGX agent

arXiv:2605.10810v1 Announce Type: new Abstract: We introduce an automatically generated benchmark for predicting hidden text in technical papers. A paper supplies visible context X and a hidden contin

24 Apr 2026

Crystal: Characterizing Relative Impact of Scholarly Publications

SafetyDGX agent

arXiv:2603.26791v2 Announce Type: replace-cross Abstract: Assessing a cited paper's impact is typically done by analyzing its citation context in isolation within the citing paper. While this focuses

23 Apr 2026

OpenCLAW-P2P v6.0: Resilient Multi-Layer Persistence, Live Reference Verification, and Production-Scale Evaluation of Decentralized AI Peer Review

HardwareDGX agent

arXiv:2604.19792v1 Announce Type: new Abstract: This paper presents OpenCLAW-P2P v6.0, a comprehensive evolution of the decentralized collective-intelligence platform in which autonomous AI agents pub

11 Apr 2026

NVIDIA’s New AI Shouldn’t Work…But It Does

HardwareDGX agent

This is a Two Minute Papers episode by Dr. Károly Zsolnai-Fehér reviewing a counterintuitive NVIDIA AI research result — likely covering a technique that defies conventional expectations yet delivers

3 Aug 2026

Can AI Evaluate AI Scientists? A Benchmarking Study of Autonomous Research Generation Systems Using Automated Multi-Model Review

Model ReleasesDGX agent

arXiv:2607.28631v1 Announce Type: new Abstract: AI Scientist systems capable of autonomous research have the potential to significantly accelerate scientific discovery. However, evaluating and compari

23 Jul 2026

I trained a 0.5M model on 1B tokens of Fineweb-edu dataset.

Model ReleasesDGX agent

Hi everyone, About a month ago I publish my very first research paper on my neural network architecture called Silia. You can look at the model here: https://huggingface.co/Srijan-Srivastava/Silia-v2

7 Jul 2026

A Vision Based System for Guided and Collaborative Reconstruction of Fragmented Documents

ResearchDGX agent

arXiv:2607.03621v1 Announce Type: new Abstract: This paper presents the development and evaluation of a collaborative system for real-time reconstruction of fragmented paper documents in the context o

GRASP: Graph-Reasoning Aided Survey Planning for High-Fidelity Related Work Generation

ResearchDGX agent

arXiv:2607.03709v1 Announce Type: new Abstract: Writing a literature review requires a deep understanding of the relationships among cited papers: how they build on, challenge, or offer alternative pe

3 Jul 2026

Game Physics Just Got 170 Times Faster

ResearchDGX agent

This video from Two Minute Papers discusses groundbreaking research by NVIDIA that achieves a 100x speedup in physics-based character animations using AI-powered super-resolution applied to physics si

7 May 2026

Stop Automating Peer Review Without Rigorous Evaluation

ResearchDGX agent

arXiv:2605.03202v1 Announce Type: new Abstract: Large language models offer a tempting solution to address the peer review crisis. This position paper argues that today's AI systems should not be used

11 Aug 2026

Stealing Reasoning Traces from Proprietary LLM APIs

Model ReleasesDGX agent

Stealing Reasoning Traces from Proprietary LLM APIs A vanity domain name (stolen-thoughts.com) for a neat paper: Anthropic, OpenAI, and Google return encrypted chain-of-thought blocks to clients that

10 Aug 2026

HNR-DAC: Hard-Negative Reranking and Distribution-Aligned Classification for Scientific Claim Verification

ResearchDGX agent

arXiv:2608.07204v1 Announce Type: new Abstract: Scientific claim verification over a cited paper requires predicting the claim--paper relation and identifying the paragraphs that justify that predicti

24 Jul 2026

Can an AI System Be Creative? A Critical Perspective from Art and Engineering

ResearchDGX agent

arXiv:2607.20796v1 Announce Type: new Abstract: This paper examines the question of whether artificial intelligence (AI) systems can be creative, approached from the dual perspective of a researcher t

22 Jul 2026

OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened

Model ReleasesDGX agent

This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke

21 Jul 2026

Very cool idea to convert memory to skills. (bookmark it) Most agent memory systems retrieve past traces as passive context. MSCE turns them…

AgentsDGX agent

Very cool idea to convert memory to skills. (bookmark it) Most agent memory systems retrieve past traces as passive context. MSCE turns them into executable skills instead. The training-free framework

16 Jul 2026

Learning Engagement Assistant (LEA): Cross-Course Scalability and Classroom Evaluation of an Agentic AI Tutoring System

AgentsDGX agent

arXiv:2607.13370v1 Announce Type: cross Abstract: This paper is an extension of a paper presented at the ICAART 2026 conference, which introduced LEA (Learning Engagement Assistant), an adaptive AI tu

15 Jul 2026

Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction

SafetyDGX agent

arXiv:2607.12835v1 Announce Type: new Abstract: Rubric-based evaluation is a promising approach for assessing open-ended outputs from LLM-based research agents, particularly in paper reproduction, whe

IdeaTrail: Full-Process Agent Trajectories for Scientific Ideation

AgentsDGX agent

arXiv:2607.10144v2 Announce Type: replace Abstract: Scientific ideation unfolds over multiple stages, including literature search, paper reading, tool use, claim checking, cross-paper synthesis, brain

Wiki Lint Report — 2026-07-15

SynthesesDGX agent

Automated lint: 26 errors, 6728 warnings, 3 info

Atomic Units of X: The Compression Layer of Intelligence

ApplicationsDGX agent

arXiv:2607.12634v1 Announce Type: new Abstract: This paper proposes a theoretical framework for understanding intelligence as a process of atomic compression and compositional reuse. We argue that cog

2 Jul 2026

Phantom References: Hallucinated Citations That Survive Peer Review at Top-Tier Conferences

ResearchDGX agent

arXiv:2607.00738v1 Announce Type: cross Abstract: Large language models can generate polished scientific text that includes unsupported claims, allowing hallucinations to enter the archival record. As

6 Jun 2026

Towards AI epidemiology: a measurement standardisation framework for prospective risk detection

SafetyDGX agent

arXiv:2512.15783v3 Announce Type: replace Abstract: This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospectiv

31 May 2026

When are ICML openreviews made public? [R]

ResearchDGX agent

Reviews and discussions for all accepted papers at ICML are made public on OpenReview after the reviewing period concludes. Authors of rejected papers may also opt-in to have their reviews and discuss

13 May 2026

The power of LLMs on your data, more than two orders of magnitude faster and cheaper

Model ReleasesDGX agent

Databases have introduced new AI-powered SQL functions which take natural language instructions as input and are evaluated using LLMs. They leverage the power of LLMs to answer new kinds of queries: W

27 Apr 2026

Pay attention to this one, AI devs. If you're building multi-agent systems, you're probably wiring static org charts. New research argues th…

AgentsDGX agent

Pay attention to this one, AI devs. If you're building multi-agent systems, you're probably wiring static org charts. New research argues they should look more like a labor market. The paper introduce

25 Apr 2026

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our…

Model ReleasesDGX agent

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our new paper: “TRINITY: An Evolved LLM Coordinator”, published

20 Apr 2026

AISysRev -- LLM-based Tool for Title-abstract Screening

Model ReleasesDGX agent

arXiv:2510.06708v3 Announce Type: replace-cross Abstract: Conducting systematic reviews is laborious. In the screening or study selection phase, the number of papers can be overwhelming. Recent resear

1 Aug 2026

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, …

Model ReleasesDGX agent

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, 17 real tasks from 3 repositories, with context-injection st

31 Jul 2026

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

Model ReleasesDGX agent

arXiv:2607.28618v1 Announce Type: new Abstract: Chemistry literature synthesis often requires assembling specific findings scattered across many publications, yet existing literature-search systems pr

30 Jul 2026

SciFigAlign: Scoring Scientific Figures by Fine-tuned Alignment of Visuals with Manuscript Evidence

SafetyDGX agent

arXiv:2607.27066v1 Announce Type: new Abstract: Scientific figure assessment in peer review differs fundamentally from general image quality evaluation: a figure must be visually legible, faithfully s

28 Jul 2026

NEO: NeRF It Once, Edit It Many Times for Continuous Object Manipulation

Model ReleasesDGX agent

arXiv:2607.24538v1 Announce Type: new Abstract: In this paper, we present NEO, a unified framework providing language-guided NeRF editing for robotic manipulation. Our paper introduces (i) a language-

8 Jul 2026

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

Model ReleasesDGX agent

arXiv:2607.05743v1 Announce Type: cross Abstract: AI coding agents now read repositories, call tools, and execute shell commands with limited human oversight, and a fast-growing body of work studies w

← Previous
1…56789…201
Next →