AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

What Do AI Agents Actually Change? An Empirical Taxonomy of Mutation Patterns in Performance-Improving Pull Requests

DGX agent

arXiv:2607.05666v1 Announce Type: cross Abstract: AI coding agents are black boxes: we cannot inspect how they generate code, but we can inspect what they change. This distinction matters for search-b

agentsarxiv-cs-ai
8 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Agentic-V2X: Small Language Model Agents for Deadline-Aware V2X Scheduling in 5G/6G Networks

DGX agent

arXiv:2607.04290v1 Announce Type: cross Abstract: Large Language Models (LLMs) are proposed as control interfaces for next-generation networks, but their latency, hallucinations, and lack of control g

local-aiarxiv-cs-ai
7 Jul 2026
Agents

Compressing the Validation Bottleneck: An Agentic Self-Driving Lab for Scientific Discovery

DGX agent

arXiv:2607.04508v1 Announce Type: new Abstract: Agentic AI-for-Science can automate ideation, planning, and analysis, but final validation still depends on real experiments. A self-driving lab (SDL) c

agentsarxiv-cs-ai
7 Jul 2026
Agents

Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration

DGX agent

arXiv:2509.10656v2 Announce Type: replace-cross Abstract: For groups of autonomous agents to achieve a particular goal, they must engage in coordination and long-horizon reasoning. Rather than relying

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Diverse Evidence, Better Forecasts: Multi-Agent Deliberation Under Information Asymmetry

DGX agent

arXiv:2607.01661v1 Announce Type: new Abstract: Multi-agent systems are increasingly used for forecasting future events, as deliberation among multiple LLMs is believed to improve reasoning and calibr

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Janus: a Playground for User-Involved Agentic Permission Management

DGX agent

arXiv:2607.01510v1 Announce Type: new Abstract: AI agents that autonomously execute tool calls on a user's behalf raise pressing questions about permission management: what role could users play, and

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

Same-Origin Policy for Agentic Browsers

DGX agent

arXiv:2606.14027v3 Announce Type: replace-cross Abstract: Agentic browsers integrate autonomous AI agents into web browsers, enabling users to accomplish web tasks through natural-language instruction

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

CAPTCHA Solving for Native GUI Agents: Automated Reasoning-Action Data Generation and Self-Corrective Training

DGX agent

arXiv:2603.23559v2 Announce Type: replace-cross Abstract: GUI agents are rapidly shifting from multi-module pipelines to end-to-end, native vision-language models (VLMs) that perceive raw screenshots

agentsarxiv-cs-ai
30 Jun 2026
Agents

NaLA: A 3D Native LLM Layout Agent for High-quality 3D Scene Generation

DGX agent

arXiv:2606.29395v1 Announce Type: new Abstract: Recently, Large Language Models (LLMs) have emerged as promising layout agents for 3D scene generation. Existing layout agents still suffer from implaus

agentsarxiv-cs-cv
30 Jun 2026
Model Releases

S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence

DGX agent

arXiv:2606.20515v2 Announce Type: replace Abstract: Real-world spatial intelligence requires reasoning over a continuous and evolving 3D world, yet existing VLMs and tool-augmented agents largely rema

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

SWE-Together: Evaluating Coding Agents in Interactive User Sessions

DGX agent

arXiv:2606.29957v1 Announce Type: cross Abstract: Most coding-agent benchmarks are static: an agent receives a complete task description up front and is judged only by its final code. Real coding assi

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Grounded Iterative Language Planning: How Parameterized World Models Reduce Hallucination Propagation in LLM Agents

DGX agent

arXiv:2606.27806v1 Announce Type: new Abstract: World models for language agents come in two useful forms. An agent-based world model calls an LLM API and reasons flexibly in language, but its errors

agentsarxiv-cs-ai
29 Jun 2026
Agents

HiLSVA: Design and Evaluation of a Human-in-the-Loop Agentic System for Scientific Visualization

DGX agent

arXiv:2606.26614v1 Announce Type: cross Abstract: Large language model (LLM) agents enable natural language interaction for scientific visualization (SciVis). Still, prior systems have essentially pri

agentsarxiv-cs-ai
26 Jun 2026
Agents

NeuraDock Visual Cognitive Load Agent Tutorial: A Quality-Gated Open-Source EEG Workflow for Alpha Dynamics and Real-Time Applications

DGX agent

arXiv:2606.26518v1 Announce Type: new Abstract: This tutorial paper provides a step-by-step, reproducible walkthrough of NeuraDock Agent, an open-source EEG agent focused on Alpha dynamics and visual

agentsarxiv-cs-ai
26 Jun 2026
Model Releases

Agent-as-a-Router: Agentic Model Routing for Coding Tasks

DGX agent

arXiv:2606.22902v2 Announce Type: replace Abstract: Real-world users typically have access to multiple Large Language Models (LLMs) from different providers, and these LLMs often excel at distinct dom

model-releasesarxiv-cs-ai
25 Jun 2026
Agents

Cooperative-ORCA*: Real-Time Proactive Deadlock Avoidance for Continuous-Space Multi-Agent Navigation

DGX agent

arXiv:2606.22757v1 Announce Type: new Abstract: Multi-Agent Path Finding (MAPF) is a problem that requires computing collision-free paths for a set of agents from their start locations to designated g

agentsarxiv-cs-ro
23 Jun 2026
Model Releases

Exploration Structure in LLM Agents for Multi-File Change Localization

DGX agent

arXiv:2606.11976v1 Announce Type: cross Abstract: Software engineering tools increasingly rely on LLM based agents to localize files to change to resolve a software issue. Most AI agents explore repos

model-releasesarxiv-cs-ai
11 Jun 2026
Agents

SkillJuror: Measuring How Agent Skill Organization Changes Runtime Behavior

DGX agent

arXiv:2606.11543v1 Announce Type: new Abstract: Agent Skills augment large language model (LLM) agents with procedural knowledge at inference time, but current benchmarks rarely distinguish what a Ski

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

CollabSkill: Evaluating Human-Agent Collaboration On Real-World Tasks

DGX agent

arXiv:2606.09833v1 Announce Type: cross Abstract: AI agents are reshaping the workspace, leading to drastic change of how humans work. Despite the considerable potential of human-agent collaboration b

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

Data Journalist Agent: Transforming Data into Verifiable Multimodal Stories

DGX agent

arXiv:2606.11176v1 Announce Type: cross Abstract: Data tells stories that shape society; the data journalist's job is to turn raw information into stories non-experts can trust. A high-quality news fe

agentsarxiv-cs-cl
10 Jun 2026
Model Releases

Frontier Coding Agents Use Metaprogramming to Adapt to Unfamiliar Programming Languages

DGX agent

arXiv:2606.10933v1 Announce Type: new Abstract: LLM-based coding agents are usually evaluated in familiar software settings: mainstream languages, common libraries, and public repositories. These benc

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

HIPIF: Hierarchical Planning and Information Folding for Long-Horizon LLM Agent Learning

DGX agent

arXiv:2606.10507v1 Announce Type: new Abstract: While Large Language Models (LLMs) have demonstrated strong capabilities as autonomous agents across a wide range of tasks, their performance often degr

agentsarxiv-cs-ai
10 Jun 2026
Agents

A case study of evaluating AI agents on a neuroscience data-to-discovery pipeline

DGX agent

arXiv:2606.07718v1 Announce Type: new Abstract: Agentic AI tools offer a promising path to automating software development bottlenecks in scientific research pipelines, particularly for stages that ta

agentsarxiv-cs-ai
9 Jun 2026
Model Releases

Executable World Models for ARC-AGI-3 in the Era of Coding Agents

DGX agent

arXiv:2605.05138v2 Announce Type: replace Abstract: We evaluate an initial coding-agent system for ARC-AGI-3 in which the agent maintains an executable Python world model, verifies it against previous

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

MemToolAgent overview with a simple restaurant booking scenario where the agent retrieves similar memories, receives feedback on an invalid time format, and generates a reflection to update its memory

DGX agent

arXiv:2606.07909v1 Announce Type: new Abstract: Modern large language model (LLM) agents can use external tools to help users solve complex tasks. However, for problems that require learning from long

agentsarxiv-cs-ai
9 Jun 2026
Agents

DuMate-DeepResearch: An Auditable Multi-Agent System with Recursive Search and Rubric-Grounded Reasoning

DGX agent

arXiv:2606.07299v1 Announce Type: new Abstract: Deep Research (DR) has emerged as a new agentic paradigm to tackle complex, open-ended research tasks, demanding systems that can iteratively frame prob

agentsarxiv-cs-ai
8 Jun 2026
Agents

Measuring Agents in Production

DGX agent

arXiv:2512.04123v4 Announce Type: replace-cross Abstract: LLM-based agents already operate in production across many industries, yet we lack an understanding of what technical methods make deployments

agentsarxiv-cs-ai
8 Jun 2026
Agents

Beyond Similarity: Trustworthy Memory Search for Personal AI Agents

DGX agent

arXiv:2606.06054v1 Announce Type: new Abstract: Personal AI agents increasingly rely on long-term memory to provide persistent personalization across sessions. However, existing memory pipelines are l

agentsarxiv-cs-ai
6 Jun 2026
Model Releases

SMAC-Talk: A Natural Language Extension of the StarCraft Multi-Agent Challenge for Large Language Models

DGX agent

arXiv:2606.04202v1 Announce Type: new Abstract: As LLMs become more widely deployed, they are increasingly expected to work alongside other AI agents rather than operating in isolation. Effective coor

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

EvoDS: Self-Evolving Autonomous Data Science Agent with Skill Learning and Context Management

DGX agent

arXiv:2606.03841v1 Announce Type: new Abstract: Recent progress in Large Language Model (LLM) agents has enabled promising advances in automated data science. However, existing approaches remain funda

agentsarxiv-cs-ai
3 Jun 2026
Safety

LAP: An Agent-to-Instrument Protocol for Autonomous Science

DGX agent

arXiv:2606.03755v1 Announce Type: new Abstract: Autonomous science is moving from demonstration to infrastructure. Large language model agents now plan experiments, and self-driving laboratories execu

safetyarxiv-cs-ai
3 Jun 2026
Agents

The Deliberative Illusion: Diagnosing Factual Attrition and Stance Homogenization in Multi-Agent LLM Deliberation

DGX agent

arXiv:2606.03032v1 Announce Type: new Abstract: Multi-agent LLM systems often treat consensus as evidence of successful interaction. For deliberative problems, however, reliability depends on whether

agentsarxiv-cs-cl
3 Jun 2026
Agents

Constitutional Black-Box Monitoring for Scheming in LLM Agents

DGX agent

arXiv:2603.00829v2 Announce Type: replace-cross Abstract: Safe deployment of Large Language Model (LLM) agents in autonomous settings requires reliable oversight mechanisms. A central challenge is det

agentsarxiv-cs-ai
2 Jun 2026
Agents

Coordination Graphs for Constrained Multi-Agent Reinforcement Learning

DGX agent

arXiv:2606.02337v1 Announce Type: new Abstract: Constrained Multi-agent reinforcement learning (CMARL) faces two intertwined challenges: the joint action space grows exponentially with the number of a

agentsarxiv-cs-ai
2 Jun 2026
Agents

Deliberative Curation: A Protocol for Multi-Agent Knowledge Bases

DGX agent

arXiv:2606.00007v1 Announce Type: new Abstract: As AI agents transition from isolated tools to collaborative participants in shared knowledge ecosystems, governing collective knowledge curation become

agentsarxiv-cs-ai
2 Jun 2026
Agents

Latent Collaboration in Multi-Agent Systems

DGX agent

arXiv:2511.20639v3 Announce Type: replace-cross Abstract: Multi-agent systems (MAS) extend large language models (LLMs) from independent single-model reasoning to coordinative system-level intelligenc

agentsarxiv-cs-ai
2 Jun 2026
Agents

Not All Flips Are Conformity: Decomposing Stance Convergence in Multi-Agent LLM Debate

DGX agent

arXiv:2606.00820v1 Announce Type: new Abstract: Multi-agent debate (MAD) is a promising strategy for improving LLM reasoning, but when agents converge on a shared answer, it is unclear whether that co

agentsarxiv-cs-cl
2 Jun 2026
Agents

On Information Self-Locking in Reinforcement Learning for Active Reasoning of LLM agents

DGX agent

arXiv:2603.12109v2 Announce Type: replace Abstract: Reinforcement learning (RL) has become a de facto paradigm for building LLM-based agents that act, interact, and reason over extended task horizons.

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

PrivacyPeek: Auditing What LLM-Based Agents Acquire, Not Just What They Say

DGX agent

arXiv:2606.00152v1 Announce Type: cross Abstract: LLM-based agents are rapidly advancing, autonomously invoking external tools to complete multi-step tasks for users. However, agents often acquire mor

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

'Skill issues'': data-centric optimization of lakehouse agents

DGX agent

arXiv:2606.01185v1 Announce Type: new Abstract: Coding agents are becoming users of data infrastructure, but their success depends not only on model quality: it also depends on the skills and environm

agentsarxiv-cs-ai
2 Jun 2026
Agents

SkillRevise: Improving LLM-Authored Agent Skills via Trace-Conditioned Skill Revision

DGX agent

arXiv:2606.01139v1 Announce Type: new Abstract: Agent skills are procedural artifacts that enable LLM agents to execute workflows, verify constraints, and recover from failures. Existing self-evolving

agentsarxiv-cs-ai
2 Jun 2026
Agents

A Behavioural and Representational Evaluation of Goal-Directedness in Language Model Agents

DGX agent

arXiv:2602.08964v2 Announce Type: replace-cross Abstract: Understanding an agent's goals helps explain and predict its behaviour, yet there is no established methodology for reliably attributing goals

agentsarxiv-cs-ai
1 Jun 2026
Safety

Scalable Constrained Multi-Agent Reinforcement Learning via State Augmentation and Consensus for Separable Dynamics

DGX agent

arXiv:2605.30461v1 Announce Type: cross Abstract: We present a distributed approach for constrained Multi-Agent Reinforcement Learning (MARL) that combines state-augmented policy learning with distrib

safetyarxiv-cs-ai
1 Jun 2026
Agents

Beyond Consensus: Trace-Level Synthesis in Mixture of Agents

DGX agent

arXiv:2605.29116v1 Announce Type: new Abstract: When multiple LLM agents solve the same problem, standard practice compresses each agent's reasoning into a majority vote or layered synthesis, treating

agentsarxiv-cs-ai
29 May 2026
Safety

V2XCrafter: Learning to Generate Driving Scene Across Agents

DGX agent

arXiv:2605.29471v1 Announce Type: new Abstract: Collaborative driving systems leverage vehicle-to-everything (V2X) communication for multi-agent collaborative perception to enhance driving safety, yet

safetyarxiv-cs-cv
29 May 2026
Agents

WorldMemArena: Evaluating Multimodal Agent Memory Through Action-World Interaction

DGX agent

arXiv:2605.29341v1 Announce Type: cross Abstract: Multimodal large language models are increasingly deployed as long-horizon agents, where memory must do more than recall: it must track an evolving wo

agentsarxiv-cs-cl
29 May 2026
Safety

Agent Explorative Policy Optimization for Multimodal Agentic Reasoning

DGX agent

arXiv:2605.28774v1 Announce Type: new Abstract: Vision-language models with extended reasoning succeed on complex problems, but many real-world problems require external tools that internal reasoning

safetyarxiv-cs-cl
28 May 2026
Agents

AutoScientists: Self-Organizing Agent Teams for Long-Running Scientific Experimentation

DGX agent

arXiv:2605.28655v1 Announce Type: new Abstract: Scientific research proceeds through iterative cycles of hypothesis generation, experiment design, execution, and revision. AI agents can automate parts

agentsarxiv-cs-ai
28 May 2026
← Previous
1…2122232425…233
Next →