AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Agents

Tracking the Behavioral Trajectories of Adapting Agents

DGX agent

arXiv:2606.02536v1 Announce Type: new Abstract: Text files such as skill files, memory files, and behavioral configuration files play a central role in defining how modern agents act. Through edits by

agentsarxiv-cs-ai
2 Jun 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Dreaming Of Others: Latent Teammate Modeling In World Models For Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.31361v1 Announce Type: cross Abstract: In cooperative multi-agent reinforcement learning (MARL), agents must coordinate with partners whose internal policies and intentions are not directly

agentsarxiv-cs-ai
1 Jun 2026
Agents

Learning to Adapt: Self-Improving Web Agent via Cognitive-Aware Exploration

DGX agent

arXiv:2605.31365v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) have led to promising progress in web agents. However, existing web agents often rely on han

agentsarxiv-cs-ai
1 Jun 2026
Agents

E-valuator: Reliable Agent Verifiers with Sequential Hypothesis Testing

DGX agent

arXiv:2512.03109v2 Announce Type: replace-cross Abstract: Agentic AI systems execute a sequence of actions, such as reasoning steps or tool calls, in response to a user prompt. To evaluate the success

agentsarxiv-cs-ai
29 May 2026
Agents

Enhancing Multi-Agent Communication through Attention Steering with Context Relevance

DGX agent

arXiv:2605.30136v1 Announce Type: new Abstract: LLM-based multi-agent systems have demonstrated remarkable performance on complex tasks through collaborative reasoning. However, these systems tend to

agentsarxiv-cs-ai
29 May 2026
Agents

MINDGAMES: A Live Arena for Evaluating Social and Strategic Reasoning in Multi-Agent LLMs

DGX agent

arXiv:2605.29512v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as interactive agents, yet their capacity for social and strategic reasoning over extended intera

agentsarxiv-cs-ai
29 May 2026
Safety

A Policy-Driven Runtime Layer for Agentic LLM Serving

DGX agent

arXiv:2605.27744v1 Announce Type: new Abstract: Multi-agent LLM systems have become the dominant production workload, but the serving stack was not built for them. The agent framework above knows agen

safetyarxiv-cs-ai
28 May 2026
Agents

Agentic Literacy Debt: A Structural Problem the AI Literacy Field Has Not Yet Named

DGX agent

arXiv:2605.27396v1 Announce Type: cross Abstract: Autonomous AI agents now plan, decide, and act on behalf of users across healthcare, financial services, and workplace contexts, often without step-by

agentsarxiv-cs-ai
28 May 2026
Agents

Cyclical Entropy Eruption: Entropy Dynamics in Agent Reinforcement Learning

DGX agent

arXiv:2605.27954v1 Announce Type: new Abstract: Agentic large language models are increasingly used to solve real-world tasks by reasoning over goals, invoking tools, and interacting with external env

agentsarxiv-cs-lg
28 May 2026
Agents

Defending LLM-based Multi-Agent Systems Against Cooperative Attacks with Sentence-Level Rectification

DGX agent

arXiv:2605.28104v1 Announce Type: new Abstract: Recent years have witnessed the rapid development of Large Language Model-based Multi-Agent Systems (MAS), which excel at collaborative decision-making

agentsarxiv-cs-ai
28 May 2026
Safety

LACUNA: Safe Agents as Recursive Program Holes

DGX agent

arXiv:2605.28617v1 Announce Type: new Abstract: LLM agents increasingly act by writing code, yet a split persists between the runtime that drives the agent and the code the model writes. The runtime o

safetyarxiv-cs-ai
28 May 2026
Agents

Understanding Automated Program Repair Agents Through the Lens of Traceability: An Empirical Study

DGX agent

arXiv:2506.08311v2 Announce Type: replace-cross Abstract: Automated Program Repair (APR) agents leverage Large Language Models (LLMs) to autonomously diagnose and fix software bugs through reasoning,

agentsarxiv-cs-ai
28 May 2026
Agents

CyberEvolver: Structured Self-Evolution for Cybersecurity Agents On the Fly

DGX agent

arXiv:2605.26195v1 Announce Type: cross Abstract: LLM-based agents are increasingly used for cybersecurity tasks, but most existing systems rely on fixed, human-designed scaffolds that struggle to ada

agentsarxiv-cs-ai
27 May 2026
Agents

Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments

DGX agent

arXiv:2605.27209v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have facilitated the widespread deployment of LLMs as interactive agents capable of reasoning, planning,

agentsarxiv-cs-ai
27 May 2026
Agents

The Necessity of a Unified Framework for LLM-Based Agent Evaluation

DGX agent

arXiv:2602.03238v2 Announce Type: replace Abstract: With the advent of Large Language Models (LLMs), general-purpose agents have seen fundamental advancements. However, evaluating these agents present

agentsarxiv-cs-ai
27 May 2026
Model Releases

AMA-Bench: Evaluating Long-Horizon Memory for Agentic Applications

DGX agent

arXiv:2602.22769v3 Announce Type: replace Abstract: Large Language Models (LLMs) are deployed as autonomous agents in increasingly complex applications, where enabling long-horizon memory is critical

model-releasesarxiv-cs-ai
26 May 2026
Safety

Multi-Agent Systems are Mixtures of Experts: Who Becomes an Influencer?

DGX agent

arXiv:2605.25929v1 Announce Type: cross Abstract: The effectiveness of multi-agent LLM deliberation depends not only on the agents' individual predictions, but also on how they communicate and collabo

safetyarxiv-cs-lg
26 May 2026
Safety

Strat-Reasoner: Reinforcing Strategic Reasoning of LLMs in Multi-Agent Games

DGX agent

arXiv:2605.04906v2 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in certain reasoning tasks, they struggle in multi-agent games where the final outcome depends on the joint

safetyarxiv-cs-ai
26 May 2026
Safety

MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems

DGX agent

arXiv:2602.04431v2 Announce Type: replace Abstract: LLM-based multi-agent systems have demonstrated impressive capabilities, but they also introduce significant safety risks when individual agents fai

safetyarxiv-cs-lg
25 May 2026
Agents

LCGuard: Latent Communication Guard for Safe KV Sharing in Multi-Agent Systems

DGX agent

arXiv:2605.22786v1 Announce Type: cross Abstract: Large language model (LLM)-based multi-agent systems increasingly rely on intermediate communication to coordinate complex tasks. While most existing

agentsarxiv-cs-lg
23 May 2026
Agents

Agentic Agile-V: From Vibe Coding to Verified Engineering in Software and Hardware Development

DGX agent

arXiv:2605.20456v1 Announce Type: cross Abstract: Agentic AI coding systems can inspect repositories, plan implementation steps, edit files, call tools, run tests, and submit pull requests. These capa

agentsarxiv-cs-ai
22 May 2026
Agents

SynAE: A Framework for Measuring the Quality of Synthetic Data for Tool-Calling Agent Evaluations

DGX agent

arXiv:2605.22564v1 Announce Type: new Abstract: Today, tool-calling agents are commonly evaluated or tested on static datasets of execution traces, including input commands, agent responses, and assoc

agentsarxiv-cs-cl
22 May 2026
Agents

Learning Incentive Structures for Cooperative Resilience in Multi-Agent Systems under Social Dilemmas

DGX agent

arXiv:2601.22292v2 Announce Type: replace-cross Abstract: Multi-agent social dilemmas, such as the tragedy of the commons, capture settings where individual incentives conflict with collective well-be

agentsarxiv-cs-lg
21 May 2026
Agents

Agent Security is a Systems Problem

DGX agent

arXiv:2605.18991v1 Announce Type: cross Abstract: We take the position that agent security must be approached as a systems problem: the AI model powering the agent must be treated as an untrusted comp

agentsarxiv-cs-ai
20 May 2026
Agents

Toward Training Superintelligent Software Agents through Self-Play SWE-RL

DGX agent

arXiv:2512.18552v2 Announce Type: replace-cross Abstract: While current software agents powered by large language models (LLMs) and agentic reinforcement learning (RL) can boost programmer productivit

agentsarxiv-cs-ai
20 May 2026
Agents

CORAL: Towards Autonomous Multi-Agent Evolution for Open-Ended Discovery

DGX agent

arXiv:2604.01658v2 Announce Type: replace Abstract: Large language model (LLM)-based evolution is a promising approach for open-ended discovery, where progress requires sustained search and knowledge

agentsarxiv-cs-ai
19 May 2026
Model Releases

FML-bench: A Controlled Study of AI Research Agent Strategies from the Perspective of Search Dynamics

DGX agent

arXiv:2605.17373v1 Announce Type: cross Abstract: AI research agents accelerate ML research by automating hypothesis generation, experimentation, and empirical refinement. Existing agent strategies ra

model-releasesarxiv-cs-ai
19 May 2026
Safety

From Demographics to Survey Anchors: Evaluating LLM Agents for Modeling Retirement Attitudes

DGX agent

arXiv:2605.16303v1 Announce Type: cross Abstract: Large language models (LLM) agents may offer tools to predict human responses to surveys. A common technique for defining these agents uses only demog

safetyarxiv-cs-ai
19 May 2026
Agents

S-Bus: Automatic Read-Set Reconstruction for Multi-Agent LLM State Coordination

DGX agent

arXiv:2605.17076v1 Announce Type: cross Abstract: Concurrent LLM agents sharing mutable natural-language state produce Structural Race Conditions (SRCs): write-write and cross-shard stale-read conflic

agentsarxiv-cs-ai
19 May 2026
Agents

Skim: Speculative Execution for Fast and Efficient Web Agents

DGX agent

arXiv:2605.16565v1 Announce Type: new Abstract: Skim is a speculative execution framework for web agents that exploits the predictable structure of purpose-built websites. Today's web-agent expense is

agentsarxiv-cs-ai
19 May 2026
Agents

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

DGX agent

arXiv:2605.15301v1 Announce Type: new Abstract: Large language models (LLMs) still struggle with the rigorous reasoning demands of hard competitive programming. While recent multi-agent frameworks att

agentsarxiv-cs-ai
18 May 2026
Model Releases

Cattle Trade: A Multi-Agent Benchmark for LLM Bluffing, Bidding, and Bargaining

DGX agent

arXiv:2605.14537v1 Announce Type: new Abstract: We introduce extsc{Cattle Trade, a multi-agent benchmark for evaluating large language models (LLMs) as agents in strategic reasoning under imperfect in

model-releasesarxiv-cs-ai
15 May 2026
Agents

MetaAgent-X : Breaking the Ceiling of Automatic Multi-Agent Systems via End-to-End Reinforcement Learning

DGX agent

arXiv:2605.14212v1 Announce Type: new Abstract: Automatic multi-agent systems aim to instantiate agent workflows without relying on manually designed or fixed orchestration. However, existing automati

agentsarxiv-cs-ai
15 May 2026
Agents

Harnessing Agentic Evolution

DGX agent

arXiv:2605.13821v1 Announce Type: new Abstract: Agentic evolution has emerged as a powerful paradigm for improving programs, workflows, and scientific solutions by iteratively generating candidates, e

agentsarxiv-cs-ai
14 May 2026
Safety

Events as Triggers for Behavioral Diversity in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.12388v1 Announce Type: cross Abstract: Effective multi-agent cooperation requires agents to adopt diverse behaviors as task conditions evolve-and to do so at the right moment. Yet, current

safetyarxiv-cs-lg
13 May 2026
Agents

A Prompt-Aware Structuring Framework for Reliable Reuse of AI-Generated Content in the Agentic Web

DGX agent

arXiv:2605.09283v1 Announce Type: new Abstract: The evolution of Large Language Models (LLMs) and the software agents built on them (AI agents) marks a turning point in the transition from a human-cen

agentsarxiv-cs-ai
12 May 2026
Model Releases

An Empirical Study of Multi-Agent Collaboration for Automated Research

DGX agent

arXiv:2603.29632v2 Announce Type: replace-cross Abstract: As AI agents evolve, the community is rapidly shifting from single Large Language Models (LLMs) to Multi-Agent Systems (MAS) to overcome cogni

model-releasesarxiv-cs-ai
12 May 2026
Agents

Consistency as a Testable Property: Statistical Methods to Evaluate AI Agent Reliability

DGX agent

arXiv:2605.10516v1 Announce Type: new Abstract: This paper establishes a rigorous measurement science for AI agent reliability, providing a foundational framework for quantifying consistency under sem

agentsarxiv-cs-ai
12 May 2026
Agents

Generalization Bounds of Emergent Communications for Agentic AI Networking

DGX agent

arXiv:2605.08613v1 Announce Type: new Abstract: The evolution of 6G networking toward agentic AI networking (AgentNet) systems requires a shift from traditional data pipelines to task-aware, agentic A

agentsarxiv-cs-ai
12 May 2026
Model Releases

Pairwise is Not Enough: Hypergraph Neural Networks for Multi-Agent Pathfinding

DGX agent

arXiv:2602.06733v2 Announce Type: replace-cross Abstract: Multi-Agent Path Finding (MAPF) is a representative multi-agent coordination problem, where multiple agents are required to navigate to their

model-releasesarxiv-cs-ai
12 May 2026
Agents

Security Risks in Tool-Enabled AI Agents: A Systematic Analysis of Privileged Execution Environments

DGX agent

arXiv:2605.09721v1 Announce Type: cross Abstract: Tool-enabled AI agents are increasingly deployed in cloud-hosted environments and offered as services, where they perform side-effecting operations th

agentsarxiv-cs-ai
12 May 2026
Agents

Shepherd: A Runtime Substrate Empowering Meta-Agents with a Formalized Execution Trace

DGX agent

arXiv:2605.10913v1 Announce Type: new Abstract: We introduce Shepherd, a functional programming model that formalizes meta-agent operations on target agents as functions, with core operations mechaniz

agentsarxiv-cs-ai
12 May 2026
Agents

Workspace Optimization: How to Train Your Agent

DGX agent

arXiv:2605.09650v1 Announce Type: new Abstract: Modern agents built on frontier language models often cannot adapt their weights. What, then, remains trainable? We argue it is the agent's workspace, t

agentsarxiv-cs-ai
12 May 2026
Agents

Are LLM Agents Behaviorally Coherent? Latent Profiles for Social Simulation

DGX agent

arXiv:2509.03736v2 Announce Type: replace Abstract: The impressive capabilities of Large Language Models (LLMs) raise the possibility that synthetic agents can serve as substitutes for real participan

agentsarxiv-cs-ai
11 May 2026
Agents

MASPO: Joint Prompt Optimization for LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.06623v1 Announce Type: cross Abstract: Large language model (LLM)-based Multi-agent systems (MAS) have shown promise in tackling complex collaborative tasks, where agents are typically orch

agentsarxiv-cs-lg
11 May 2026
Agents

Searching for Privacy Risks in LLM Agents via Simulation

DGX agent

arXiv:2508.10880v3 Announce Type: replace-cross Abstract: The widespread deployment of LLM-based agents is likely to introduce a critical privacy threat: malicious agents that proactively engage other

agentsarxiv-cs-ai
11 May 2026
Safety

Who Prices Cognitive Labor in the Age of Agents? Compute-Anchored Wages

DGX agent

arXiv:2605.05558v2 Announce Type: replace Abstract: A natural intuition about the economics of AI agents is that, because agents can be replicated at very low marginal cost, agent labor may be supplie

safetyarxiv-cs-ai
11 May 2026
Model Releases

Terminus-4B: Can a Smaller Model Replace Frontier LLMs at Agentic Execution Tasks?

DGX agent

arXiv:2605.03195v1 Announce Type: new Abstract: Modern coding agents increasingly delegate specialized subtasks to subagents, which are smaller, focused agentic loops that handle narrow responsibiliti

model-releasesarxiv-cs-ai
7 May 2026
← Previous
1…1819202122…230
Next →