AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

PIVOT: Bridging Planning and Execution in LLM Agents via Trajectory Refinement

DGX agent

arXiv:2605.11225v1 Announce Type: cross Abstract: Large language model (LLM)-based agents frequently generate seemingly coherent plans that fail upon execution due to infeasible actions, constraint vi

agentsarxiv-cs-lg
13 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Robust Multi-Agent Path Finding under Observation Attacks: A Principled Adversarial-Plus-Smoothing Training Recipe

DGX agent

arXiv:2605.11469v1 Announce Type: new Abstract: Decentralized multi-agent path finding (MAPF) routes a team of agents on a shared grid, each acting from its own local view. The standard solution train

safetyarxiv-cs-lg
13 May 2026
Agents

ASIA: an Autonomous System Identification Agent

DGX agent

arXiv:2605.10480v1 Announce Type: new Abstract: Over the years, research in system identification has provided a rich set of methods for learning dynamical models, together with well-established theor

agentsarxiv-cs-ai
12 May 2026
Agents

Context Learning for Multi-Agent Discussion

DGX agent

arXiv:2602.02350v2 Announce Type: replace Abstract: Multi-Agent Discussion (MAD) has garnered increasing attention very recently, where multiple LLM instances collaboratively solve problems via struct

agentsarxiv-cs-ai
12 May 2026
Agents

From Spark to Fire: Modeling and Mitigating Error Cascades in LLM-Based Multi-Agent Collaboration

DGX agent

arXiv:2603.04474v2 Announce Type: replace-cross Abstract: Large Language Model-based Multi-Agent Systems (LLM-MAS) are increasingly applied to complex collaborative scenarios. However, their collabora

agentsarxiv-cs-ai
12 May 2026
Agents

HAGE: Harnessing Agentic Memory via RL-Driven Weighted Graph Evolution

DGX agent

arXiv:2605.09942v1 Announce Type: new Abstract: Memory retrieval in agentic large language model (LLM) systems is often treated as a static lookup problem, relying on flat vector search or fixed binar

agentsarxiv-cs-ai
12 May 2026
Agents

Heteroscedastic Diffusion for Multi-Agent Trajectory Modeling

DGX agent

arXiv:2605.10717v1 Announce Type: cross Abstract: Multi-agent trajectory modeling traditionally focuses on forecasting, often neglecting more general tasks like trajectory completion, which is essenti

agentsarxiv-cs-cv
12 May 2026
Safety

Iterative Critique-and-Routing Controller for Multi-Agent Systems with Heterogeneous LLMs

DGX agent

arXiv:2605.08686v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems often rely on a controller to coordinate a pool of heterogeneous models, yet existing controllers are typ

safetyarxiv-cs-ai
12 May 2026
Agents

MAGE: Multi-Agent Self-Evolution with Co-Evolutionary Knowledge Graphs

DGX agent

arXiv:2605.10064v1 Announce Type: new Abstract: Self-evolving language-model agents must decide what to learn next and how to preserve what they have learned across iterations. Existing systems typica

agentsarxiv-cs-ai
12 May 2026
Safety

Mechanism Design Is Not Enough: Prosocial Agents for Cooperative AI

DGX agent

arXiv:2605.08426v1 Announce Type: cross Abstract: Ensuring that AI agents behave safely and beneficially when interacting with other parties has emerged as one of the central challenges of modern AI s

safetyarxiv-cs-ai
12 May 2026
Agents

MemReread: Enhancing Agentic Long-Context Reasoning via Memory-Guided Rereading

DGX agent

arXiv:2605.10268v1 Announce Type: cross Abstract: To tackle long-context reasoning tasks without the quadratic complexity of standard attention mechanisms, approaches based on agent memory have emerge

agentsarxiv-cs-ai
12 May 2026
Agents

NyayaAI: An AI-Powered Legal Assistant Using Multi-Agent Architecture and Retrieval-Augmented Generation

DGX agent

arXiv:2605.10155v1 Announce Type: new Abstract: Legal information in India remains largely inaccessible due to the complexity of legal language and the sheer volume of legal documentation involved in

agentsarxiv-cs-cl
12 May 2026
Model Releases

PRISM: Generation-Time Detection and Mitigation of Secret Leakage in Multi-Agent LLM Pipelines

DGX agent

arXiv:2605.10614v1 Announce Type: new Abstract: Multi-agent LLM systems introduce a security risk in which sensitive information accessed by one agent can propagate through shared context and reappear

model-releasesarxiv-cs-ai
12 May 2026
Agents

Remember the Decision, Not the Description: A Rate-Distortion Framework for Agent Memory

DGX agent

arXiv:2605.10870v1 Announce Type: new Abstract: Long-horizon language agents must operate under limited runtime memory, yet existing memory mechanisms often organize experience around descriptive crit

agentsarxiv-cs-ai
12 May 2026
Agents

Skill-CMIB: Multimodal Agent Skill for Consistent Action via Conditional Multimodal Information Bottleneck

DGX agent

arXiv:2605.08526v1 Announce Type: new Abstract: While LLM-based agents excel at planning and executing long action sequences, their execution often remains inconsistent across trials, limiting reliabi

agentsarxiv-cs-lg
12 May 2026
Model Releases

Towards On-Policy Data Evolution for Visual-Native Multimodal Deep Search Agents

DGX agent

arXiv:2605.10832v1 Announce Type: new Abstract: Multimodal deep search requires an agent to solve open-world problems by chaining search, tool use, and visual reasoning over evolving textual and visua

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

WildClawBench: A Benchmark for Real-World, Long-Horizon Agent Evaluation

DGX agent

arXiv:2605.10912v1 Announce Type: new Abstract: Large language and vision-language models increasingly power agents that act on a user's behalf through command-line interface (CLI) harnesses. However,

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

AgentEscapeBench: Evaluating Out-of-Domain Tool-Grounded Reasoning in LLM Agents

DGX agent

arXiv:2605.07926v1 Announce Type: new Abstract: As LLM-based agents increasingly rely on external tools, it is important to evaluate their ability to sustain tool-grounded reasoning beyond familiar wo

model-releasesarxiv-cs-ai
11 May 2026
Agents

Dual-Agent Co-Training for Health Coaching via Implicit Adversarial Preference Optimization

DGX agent

arXiv:2605.07011v1 Announce Type: new Abstract: Motivational-interviewing-based health coaching is an effective approach for improving mental health and promoting healthy behavior change. However, the

agentsarxiv-cs-lg
11 May 2026
Agents

GraphDC: A Divide-and-Conquer Multi-Agent System for Scalable Graph Algorithm Reasoning

DGX agent

arXiv:2605.06671v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated strong potential for many mathematical problems. However, their performance on graph algorithmic tasks is

agentsarxiv-cs-ai
11 May 2026
Safety

Hidden Coalitions in Multi-Agent AI: A Spectral Diagnostic from Internal Representations

DGX agent

arXiv:2605.06696v1 Announce Type: new Abstract: Collections of interacting AI agents can form coalitions, creating emergent group-level organization that is critical for AI safety and alignment. Howev

safetyarxiv-cs-ai
11 May 2026
Model Releases

Learning Agent Routing From Early Experience

DGX agent

arXiv:2605.07180v1 Announce Type: new Abstract: LLM agents achieve strong performance on complex reasoning tasks but incur high latency and compute cost. In practice, many queries fall within the capa

model-releasesarxiv-cs-cl
11 May 2026
Agents

MEMOREPAIR: Barrier-First Cascade Repair in Agentic Memory

DGX agent

arXiv:2605.07242v1 Announce Type: new Abstract: Agentic memory evolves across tasks into durable derived artifacts: summaries, cached outputs, embeddings, learned skills, and executable tool procedure

agentsarxiv-cs-ai
11 May 2026
Agents

Securing Computer-Use Agents: A Unified Architecture-Lifecycle Framework for Deployment-Grounded Reliability

DGX agent

arXiv:2605.07110v1 Announce Type: new Abstract: Computer-use agents(CUAs)are moving frombounded benchmarks toward real software environments, wherethey operate browsers, desktops, mobile applications,

agentsarxiv-cs-cl
11 May 2026
Model Releases

SREGym: A Live Benchmark for AI SRE Agents with High-Fidelity Failure Scenarios

DGX agent

arXiv:2605.07161v1 Announce Type: new Abstract: AI agents are increasingly used to diagnose and mitigate failures in production systems, known as agentic Site Reliability Engineering (SRE). Current SR

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

TeamBench: Evaluating Agent Coordination under Enforced Role Separation

DGX agent

arXiv:2605.07073v1 Announce Type: new Abstract: Agent systems often decompose a task across multiple roles, but these roles are typically specified by prompts rather than enforced by access controls.

model-releasesarxiv-cs-ai
11 May 2026
Agents

VIDEE: Visual and Interactive Decomposition, Execution, and Evaluation of Text Analytics with Intelligent Agents

DGX agent

arXiv:2506.21582v5 Announce Type: replace-cross Abstract: Text analytics has traditionally required specialized knowledge in Natural Language Processing (NLP) or text analysis, which presents a barrie

agentsarxiv-cs-ai
11 May 2026
Agents

An Agent-Oriented Pluggable Experience-RAG Skill for Experience-Driven Retrieval Strategy Orchestration

DGX agent

arXiv:2605.03989v1 Announce Type: new Abstract: Retrieval-augmented generation systems often assume that one fixed retrieval pipeline is sufficient across heterogeneous tasks, yet factoid question ans

agentsarxiv-cs-ai
7 May 2026
Agents

ARMATA: Auto-Regressive Multi-Agent Task Assignment

DGX agent

arXiv:2605.04225v1 Announce Type: cross Abstract: Coordinating multi-agent systems over spatially distributed areas requires solving a complex hierarchical problem: first distributing areas among agen

agentsarxiv-cs-ro
7 May 2026
Agents

Learning to Orchestrate Agents in Natural Language with the Conductor

DGX agent

arXiv:2512.04388v5 Announce Type: replace Abstract: Powerful large language models (LLMs) from different providers have been expensively trained and finetuned to specialize across varying domains. In

agentsarxiv-cs-lg
7 May 2026
Safety

LLM-Based Human-Agent Collaboration and Interaction Systems: A Survey

DGX agent

arXiv:2505.00753v5 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing interest in building fully autonomous agents. However, fully autonomous LLM-bas

safetyarxiv-cs-cl
7 May 2026
Agents

Self-Induced Outcome Potential: Turn-Level Credit Assignment for Agents without Verifiers

DGX agent

arXiv:2605.04984v1 Announce Type: cross Abstract: Long-horizon LLM agents depend on intermediate information-gathering turns, yet training feedback is usually observed only at the final answer, becaus

agentsarxiv-cs-cl
7 May 2026
Local Ai

An Empirical Study of Agent Skills for Healthcare: Practice, Gaps, and Governance

DGX agent

arXiv:2605.02709v1 Announce Type: new Abstract: Healthcare automation is shaped by local procedures and organizational constraints, so agent capabilities rarely transfer unchanged across settings. Age

local-aiarxiv-cs-ai
6 May 2026
Agents

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects

DGX agent

arXiv:2511.20657v2 Announce Type: replace-cross Abstract: The development of agents with emotional intelligence is becoming increasingly vital due to their significant role in human-computer interacti

agentsarxiv-cs-ai
6 May 2026
Safety

Position: Safety and Fairness in Agentic AI Depend on Interaction Topology, Not on Model Scale or Alignment

DGX agent

arXiv:2605.01147v1 Announce Type: new Abstract: As large language models are increasingly deployed as interacting agents in high-stakes decisions, the AI safety community assumes that safety propertie

safetyarxiv-cs-ai
6 May 2026
Agents

Runtime Evaluation of Procedural Content Generation in an Endless Runner Game Using Autonomous Agents

DGX agent

arXiv:2605.01783v1 Announce Type: new Abstract: Procedural Content Generation (PCG) enables game content to be created algorithmically without direct manual level-design effort, but it introduces a se

agentsarxiv-cs-ai
6 May 2026
Agents

Truth or Tribe: How In-group Favoritism Prioritize Facts in Persona Agents

DGX agent

arXiv:2605.01329v1 Announce Type: new Abstract: In-group favoritism refers to the phenomena of favoring members of one's in-group over out-group members and is widely observed in numerous social coope

agentsarxiv-cs-ai
6 May 2026
Model Releases

Workspace-Bench 1.0: Benchmarking AI Agents on Workspace Tasks with Large-Scale File Dependencies

DGX agent

arXiv:2605.03596v1 Announce Type: cross Abstract: Workspace learning requires AI agents to identify, reason over, exploit, and update explicit and implicit dependencies among heterogeneous files in a

model-releasesarxiv-cs-cl
6 May 2026
Agents

A Language for Describing Agentic LLM Contexts

DGX agent

arXiv:2605.01920v1 Announce Type: cross Abstract: Large language models are increasingly used within larger systems ('LLM agents'). These make a sequence of LLM calls, each call providing the LLM with

agentsarxiv-cs-cl
5 May 2026
Agents

TADI: Tool-Augmented Drilling Intelligence via Agentic LLM Orchestration over Heterogeneous Wellsite Data

DGX agent

arXiv:2605.00060v1 Announce Type: new Abstract: We present TADI (Tool-Augmented Drilling Intelligence), an agentic AI system that transforms drilling operational data into evidence-based analytical in

agentsarxiv-cs-ai
5 May 2026
Model Releases

Watermarking LLM Agent Trajectories

DGX agent

arXiv:2602.18700v2 Announce Type: replace-cross Abstract: LLM agents rely heavily on high-quality trajectory data to guide their problem-solving behaviors, yet producing such data requires substantial

model-releasesarxiv-cs-cl
5 May 2026
Safety

When Embedding-Based Defenses Fail: Rethinking Safety in LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.01133v1 Announce Type: cross Abstract: Large language model (LLM)-powered multi-agent systems (MAS) enable agents to communicate and share information, achieving strong performance on compl

safetyarxiv-cs-lg
5 May 2026
Agents

Affordance Agent Harness: Verification-Gated Skill Orchestration

DGX agent

arXiv:2605.00663v1 Announce Type: cross Abstract: Affordance grounding requires identifying where and how an agent should interact in open-world scenes, where actionable regions are often small, occlu

agentsarxiv-cs-cv
4 May 2026
Agents

Disentangled Control of Multi-Agent Systems

DGX agent

arXiv:2511.05900v3 Announce Type: replace-cross Abstract: This paper develops a general framework for multi-agent control synthesis, which applies to a wide range of problems with convergence guarante

agentsarxiv-cs-ro
4 May 2026
Agents

Group Cognition Learning: Making Everything Better Through Governed Two-Stage Agents Collaboration

DGX agent

arXiv:2605.00370v1 Announce Type: new Abstract: Centralized multimodal learning commonly compresses language, acoustic, and visual signals into a single fused representation for prediction. While effe

agentsarxiv-cs-lg
4 May 2026
Safety

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration

DGX agent

arXiv:2605.00444v1 Announce Type: new Abstract: Multi-modal large language models (MLLMs) advance vision language understanding but face inherent limitations in long-video tasks due to bounded percept

safetyarxiv-cs-cv
4 May 2026
Agents

Agentic Compilation: Mitigating the LLM Rerun Crisis for Minimized-Inference-Cost Web Automation

DGX agent

arXiv:2604.09718v2 Announce Type: cross Abstract: LLM-driven web agents operating through continuous inference loops -- repeatedly querying a model to evaluate browser state and select actions -- exhi

agentsarxiv-cs-ai
1 May 2026
Model Releases

AgenticRecTune: Multi-Agent with Self-Evolving Skillhub for Recommendation System Optimization

DGX agent

arXiv:2604.26969v1 Announce Type: cross Abstract: Modern large-scale recommendation systems are typically constructed as multi-stage pipelines, encompassing pre-ranking, ranking, and re-ranking phases

model-releasesarxiv-cs-ai
1 May 2026
← Previous
1…5152535455…233
Next →