AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “papers”

GridTimelineEvolution
12,059 results
Model Releases

I trained a 0.5M model on 1B tokens of Fineweb-edu dataset.

DGX agent

Hi everyone, About a month ago I publish my very first research paper on my neural network architecture called Silia. You can look at the model here: https://huggingface.co/Srijan-Srivastava/Silia-v2

model-releasesr-localllama
23 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

A Vision Based System for Guided and Collaborative Reconstruction of Fragmented Documents

DGX agent

arXiv:2607.03621v1 Announce Type: new Abstract: This paper presents the development and evaluation of a collaborative system for real-time reconstruction of fragmented paper documents in the context o

researcharxiv-cs-cv
7 Jul 2026
Research

Game Physics Just Got 170 Times Faster

DGX agent

This video from Two Minute Papers discusses groundbreaking research by NVIDIA that achieves a 100x speedup in physics-based character animations using AI-powered super-resolution applied to physics si

researchtwo-minute-papers
3 Jul 2026
Model Releases

I-WebGenBench : Evaluating Interactivity in LLM-Generated Scientific Web Applications

DGX agent

arXiv:2606.00750v1 Announce Type: new Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, existing

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

OARelatedWork: A Large-Scale Dataset of Related Work Sections with Full-texts from Open Access Sources

DGX agent

arXiv:2405.01930v2 Announce Type: replace Abstract: This paper introduces OARelatedWork: a dataset for related work generation from open-access sources. It is the first large-scale multi-document summ

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

PaperVoyager : Building Interactive Web with Visual Language Models

DGX agent

arXiv:2603.22999v3 Announce Type: replace Abstract: Recent advances in visual language models have enabled autonomous agents for complex reasoning, tool use, and document understanding. However, exist

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Likelihood scoring for continuations of mathematical text: a self-supervised benchmark with tests for shortcut vulnerabilities

DGX agent

arXiv:2605.10810v1 Announce Type: new Abstract: We introduce an automatically generated benchmark for predicting hidden text in technical papers. A paper supplies visible context X and a hidden contin

model-releasesarxiv-cs-lg
12 May 2026
Research

Stop Automating Peer Review Without Rigorous Evaluation

DGX agent

arXiv:2605.03202v1 Announce Type: new Abstract: Large language models offer a tempting solution to address the peer review crisis. This position paper argues that today's AI systems should not be used

researcharxiv-cs-ai
7 May 2026
Model Releases

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

Stealing Reasoning Traces from Proprietary LLM APIs A vanity domain name (stolen-thoughts.com) for a neat paper: Anthropic, OpenAI, and Google return encrypted chain-of-thought blocks to clients that

model-releasessimon-willison
11 Aug 2026
Research

HNR-DAC: Hard-Negative Reranking and Distribution-Aligned Classification for Scientific Claim Verification

DGX agent

arXiv:2608.07204v1 Announce Type: new Abstract: Scientific claim verification over a cited paper requires predicting the claim--paper relation and identifying the paragraphs that justify that predicti

researcharxiv-cs-cl
10 Aug 2026
Research

Can an AI System Be Creative? A Critical Perspective from Art and Engineering

DGX agent

arXiv:2607.20796v1 Announce Type: new Abstract: This paper examines the question of whether artificial intelligence (AI) systems can be creative, approached from the dual perspective of a researcher t

researcharxiv-cs-ai
24 Jul 2026
Model Releases

OpenAI’s accidental cyberattack against Hugging Face is science fiction that happened

DGX agent

This story is wild. The short version: OpenAI were running a cybersecurity test against an unreleased model, with the model's guardrail features turned off. Rather than solve the test, the model broke

model-releasessimon-willison
22 Jul 2026
Agents

Very cool idea to convert memory to skills. (bookmark it) Most agent memory systems retrieve past traces as passive context. MSCE turns them…

DGX agent

Very cool idea to convert memory to skills. (bookmark it) Most agent memory systems retrieve past traces as passive context. MSCE turns them into executable skills instead. The training-free framework

agentsdair-ai--x
21 Jul 2026
Agents

Learning Engagement Assistant (LEA): Cross-Course Scalability and Classroom Evaluation of an Agentic AI Tutoring System

DGX agent

arXiv:2607.13370v1 Announce Type: cross Abstract: This paper is an extension of a paper presented at the ICAART 2026 conference, which introduced LEA (Learning Engagement Assistant), an adaptive AI tu

agentsarxiv-cs-ai
16 Jul 2026
Safety

Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction

DGX agent

arXiv:2607.12835v1 Announce Type: new Abstract: Rubric-based evaluation is a promising approach for assessing open-ended outputs from LLM-based research agents, particularly in paper reproduction, whe

safetyarxiv-cs-cl
15 Jul 2026
Agents

IdeaTrail: Full-Process Agent Trajectories for Scientific Ideation

DGX agent

arXiv:2607.10144v2 Announce Type: replace Abstract: Scientific ideation unfolds over multiple stages, including literature search, paper reading, tool use, claim checking, cross-paper synthesis, brain

agentsarxiv-cs-ai
15 Jul 2026
Syntheses

Wiki Lint Report — 2026-07-15

DGX agent

Automated lint: 26 errors, 6728 warnings, 3 info

linthealth-checkautomated
15 Jul 2026
Research

Phantom References: Hallucinated Citations That Survive Peer Review at Top-Tier Conferences

DGX agent

arXiv:2607.00738v1 Announce Type: cross Abstract: Large language models can generate polished scientific text that includes unsupported claims, allowing hallucinations to enter the archival record. As

researcharxiv-cs-ai
2 Jul 2026
Safety

Towards AI epidemiology: a measurement standardisation framework for prospective risk detection

DGX agent

arXiv:2512.15783v3 Announce Type: replace Abstract: This paper proposes a measurement standardisation framework that compresses expert-AI interactions into structured, comparable fields for prospectiv

safetyarxiv-cs-ai
6 Jun 2026
Research

Who Annotates in NLP? A Large-scale Assessment of Human Annotation Reporting between 2018 and 2025

DGX agent

arXiv:2606.02255v1 Announce Type: cross Abstract: Human annotation is the empirical foundation of much NLP research, from dataset construction to model evaluation, but papers often leave unclear who p

researcharxiv-cs-ai
2 Jun 2026
Research

When are ICML openreviews made public? [R]

DGX agent

Reviews and discussions for all accepted papers at ICML are made public on OpenReview after the reviewing period concludes. Authors of rejected papers may also opt-in to have their reviews and discuss

researchr-machinelearning
31 May 2026
Research

LECTOR: Joint Optimization of Scientific Reasoning Graphs and Introduction Generation

DGX agent

arXiv:2605.25964v1 Announce Type: new Abstract: AI Scientists have shown promising progress across multiple stages of the research pipeline, among which automatic scientific paper writing remains a fo

researcharxiv-cs-ai
26 May 2026
Agents

Methods for Formal Verification of Agent Skills: Three Layers Toward a Mechanically Checkable Capability-Containment Proof

DGX agent

arXiv:2605.23951v1 Announce Type: new Abstract: The companion paper introduced a four-level verification lattice on agent-skill manifests (unverified, declared, tested, formal) and left the top level

agentsarxiv-cs-ai
26 May 2026
Model Releases

The power of LLMs on your data, more than two orders of magnitude faster and cheaper

DGX agent

Databases have introduced new AI-powered SQL functions which take natural language instructions as input and are evaluated using LLMs. They leverage the power of LLMs to answer new kinds of queries: W

model-releasesgoogle-cloud-ai
13 May 2026
Agents

Pay attention to this one, AI devs. If you're building multi-agent systems, you're probably wiring static org charts. New research argues th…

DGX agent

Pay attention to this one, AI devs. If you're building multi-agent systems, you're probably wiring static org charts. New research argues they should look more like a labor market. The paper introduce

agentsdair-ai--x
27 Apr 2026
Model Releases

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our…

DGX agent

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our new paper: “TRINITY: An Evolved LLM Coordinator”, published

model-releasesdavid-ha--x
25 Apr 2026
Model Releases

AISysRev -- LLM-based Tool for Title-abstract Screening

DGX agent

arXiv:2510.06708v3 Announce Type: replace-cross Abstract: Conducting systematic reviews is laborious. In the screening or study selection phase, the number of papers can be overwhelming. Recent resear

model-releasesarxiv-cs-ai
20 Apr 2026
Research

Mandatory In-Person Presentation in CVPR 2026 [D]

DGX agent

This Reddit thread on r/MachineLearning discusses CVPR 2026's policy requiring accepted papers to be registered under an in-person author registration, with virtual attendance still permitted if circu

researchr-machinelearning
13 Apr 2026
Model Releases

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences

DGX agent

arXiv:2602.11354v2 Announce Type: replace Abstract: The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, …

DGX agent

If you maintain an AGENTS.md or a CLAUDE.md, this is worth a read. (bookmark it) 288 gold-test evaluated runs across Claude Code and Codex, 17 real tasks from 3 repositories, with context-injection st

model-releasesdair-ai--x
1 Aug 2026
Model Releases

AskChem: Claim-Centered Infrastructure for Chemistry Literature Synthesis

DGX agent

arXiv:2607.28618v1 Announce Type: new Abstract: Chemistry literature synthesis often requires assembling specific findings scattered across many publications, yet existing literature-search systems pr

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

SciFigAlign: Scoring Scientific Figures by Fine-tuned Alignment of Visuals with Manuscript Evidence

DGX agent

arXiv:2607.27066v1 Announce Type: new Abstract: Scientific figure assessment in peer review differs fundamentally from general image quality evaluation: a figure must be visually legible, faithfully s

safetyarxiv-cs-cv
30 Jul 2026
Model Releases

NEO: NeRF It Once, Edit It Many Times for Continuous Object Manipulation

DGX agent

arXiv:2607.24538v1 Announce Type: new Abstract: In this paper, we present NEO, a unified framework providing language-guided NeRF editing for robotic manipulation. Our paper introduces (i) a language-

model-releasesarxiv-cs-ro
28 Jul 2026
Applications

Atomic Units of X: The Compression Layer of Intelligence

DGX agent

arXiv:2607.12634v1 Announce Type: new Abstract: This paper proposes a theoretical framework for understanding intelligence as a process of atomic compression and compositional reuse. We argue that cog

applicationsarxiv-cs-ai
15 Jul 2026
Model Releases

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

DGX agent

arXiv:2607.05743v1 Announce Type: cross Abstract: AI coding agents now read repositories, call tools, and execute shell commands with limited human oversight, and a fast-growing body of work studies w

model-releasesarxiv-cs-ai
8 Jul 2026
Research

GRASP: Graph-Reasoning Aided Survey Planning for High-Fidelity Related Work Generation

DGX agent

arXiv:2607.03709v1 Announce Type: new Abstract: Writing a literature review requires a deep understanding of the relationships among cited papers: how they build on, challenge, or offer alternative pe

researcharxiv-cs-cl
7 Jul 2026
Model Releases

PreScience: A Dataset and Benchmark for Scientific Forecasting

DGX agent

arXiv:2602.20459v2 Announce Type: replace Abstract: Can AI systems trained on the existing scientific record forecast the advances that will follow? We introduce PreScience, a dataset and benchmark fo

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

HyBIRD: Hyperbolic Bridge Retrieval and Diagnosis for Methodology Inspiration Retrieval

DGX agent

arXiv:2606.28336v1 Announce Type: cross Abstract: Methodology Inspiration Retrieval (MIR) asks a system to retrieve prior papers whose methods can inspire a new research proposal. Unlike general scien

model-releasesarxiv-cs-ai
30 Jun 2026
Research

The Governance Inversion Hypothesis: Why More AI Regulation May Produce Less Organisational Control

DGX agent

arXiv:2606.26117v1 Announce Type: cross Abstract: This paper introduces the Governance Inversion Hypothesis (GIH) to explain a growing paradox in artificial intelligence (AI) governance: under conditi

researcharxiv-cs-ai
26 Jun 2026
Research

ReviewGuard: Aligning LLM-Assisted Peer Review with Long-Term Scientific Impact

DGX agent

arXiv:2606.24892v1 Announce Type: cross Abstract: Peer review is central to scientific quality control, yet it can undervalue papers that later achieve substantial citation impact. While frontier larg

researcharxiv-cs-ai
25 Jun 2026
Agents

my weekend hobby: self improvement research

DGX agent

my weekend hobby: self improvement research in arxiv paper #2, i tackle the last topic from paper #1: @activegraphai as an architectural affordance for self-improving agents 'Regimes: An Auditable, He

agentsyohei-nakajima--x
10 Jun 2026
Applications

Merit or networks? What decides where research is published

DGX agent

arXiv:2606.03763v1 Announce Type: cross Abstract: Does scientific publishing reward the quality of ideas or the advantage of connections? The question is universal to prestige-driven science, yet it h

applicationsarxiv-cs-ai
3 Jun 2026
Safety

Review Arcade: On the Human Alignment and Gameability of LLM Reviews

DGX agent

arXiv:2605.28897v1 Announce Type: new Abstract: LLM-generated reviews for scientific papers are gaining considerable traction and are even being officially piloted by major conferences. We have to ass

safetyarxiv-cs-ai
29 May 2026
Agents

AI Research Agents Narrow Scientific Exploration

DGX agent

arXiv:2605.27905v1 Announce Type: new Abstract: AI research agents can now generate research ideas, design experiments, run code, and draft papers, raising the possibility of large-scale AI-assisted s

agentsarxiv-cs-cl
28 May 2026
Agents

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture'

DGX agent

okay i think this is a much better visualization of what i mean by 'log-centric agent architecture' babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is t

agentsyohei-nakajima--x
28 May 2026
Safety

Whose Good, Whose Place? The Moral Geography of Agentic AI for Social Good

DGX agent

arXiv:2605.22995v1 Announce Type: cross Abstract: Agentic AI systems are increasingly proposed for social-good domains, often invoking the United Nations Sustainable Development Goals (SDGs) as a voca

safetyarxiv-cs-ai
25 May 2026
Agents

recommended reading.

DGX agent

recommended reading. babyagi has ~200 citations, but 0 papers... i just published my first paper on arXiv 😆 'The Log is the Agent: Event-Sourced Reactive Graphs for Auditable, Forkable Agentic Systems

agentsyohei-nakajima--x
22 May 2026
Model Releases

No One Knows the State of the Art in Geospatial Foundation Models

DGX agent

arXiv:2605.12678v1 Announce Type: new Abstract: Geospatial foundation models (GFMs) have been proposed as generalizable backbones for disaster response, land-cover mapping, food-security monitoring, a

model-releasesarxiv-cs-cv
14 May 2026
← Previous
1…7891011…252
Next →