AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

Do Multimodal Agents Really Benefit from Tool Use? A Systematic Study of Capability Gains

DGX agent

arXiv:2606.02357v1 Announce Type: cross Abstract: Tool-augmented multimodal agents show strong benchmark gains, often taken as evidence that agents have learned to use tools. We argue that this interp

model-releasesarxiv-cs-ai
2 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Ev-Trust: An Evolutionarily Stable Trust Mechanism for Decentralized LLM-Based Multi-Agent Service Economies

DGX agent

arXiv:2512.16167v3 Announce Type: replace-cross Abstract: Decentralized LLM-based multi-agent service economies face three vulnerabilities that undermine traditional trust mechanisms: reduced cost of

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

MMG2Skill: Can Agents Distill In-the-Wild Guides into Self-Evolving Skills?

DGX agent

arXiv:2606.01993v1 Announce Type: cross Abstract: Abundant procedural knowledge on the Web holds great potential for helping agents solve long-horizon tasks. However, such knowledge is often multimoda

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

ROGUE: Misaligned Agent Behavior Arising from Ordinary Computer Use

DGX agent

arXiv:2606.00341v1 Announce Type: cross Abstract: As AI agents are increasingly deployed in real personal and corporate settings (email accounts, development workflows, company databases, etc.), safet

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

TechGraphRAG: An Agentic Graph-Augmented RAG Framework for Technical Literature Reasoning

DGX agent

arXiv:2606.01613v1 Announce Type: cross Abstract: This paper presents an agentic retrieval-augmented generation (RAG) framework for domain-specific technical reasoning support, instantiated over a cur

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

When Parallelism Pays Off: Cohesion-Aware Task Partitioning for Multi-Agent Coding

DGX agent

arXiv:2606.00953v1 Announce Type: new Abstract: Multi-agent Large Language Model (LLM) systems offer a way to decompose complex tasks, such as coding, through parallelization and context isolation. Ho

model-releasesarxiv-cs-lg
2 Jun 2026
Model Releases

Counterfactual Trace Auditing of LLM Agent Skills

DGX agent

arXiv:2605.11946v2 Announce Type: replace Abstract: Large Language Model agents are increasingly augmented with agent skills. Current evaluation methods for skills remain limited. Most deployed benchm

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

DynaTree: Dynamic Agentic Retrieval Tree for Time-Sensitive News Retrieval

DGX agent

arXiv:2605.31377v1 Announce Type: cross Abstract: Agentic Retrieval-Augmented Generation improves retrieval by integrating planning, tool use, and iterative reasoning, but existing agentic RAG methods

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

LLM Anonymization Against Agentic Re-Identificatio

DGX agent

arXiv:2605.30848v1 Announce Type: cross Abstract: Agentic LLMs with web search change the threat model for text anonymization: weak contextual cues can become cross-referenceable evidence for re-ident

agentsarxiv-cs-cl
1 Jun 2026
Agents

AgentSchool: An LLM-Powered Multi-Agent Simulation for Education

DGX agent

arXiv:2605.30144v1 Announce Type: new Abstract: Despite the rapid deployment of LLMs into classrooms, validating educational AI remains uniquely intractable: interventions act on developing learners w

agentsarxiv-cs-ai
29 May 2026
Agents

Does The Way You Plan Matter? An Empirical Study of Planning Representations for LLM Web Agents

DGX agent

arXiv:2605.29927v1 Announce Type: cross Abstract: Despite recent advances, LLM-based web agents still struggle with limited exploration, omission of critical steps, and sensitivity to task constraints

agentsarxiv-cs-ai
29 May 2026
Agents

Governing Technical Debt in Agentic AI Systems

DGX agent

arXiv:2605.29129v1 Announce Type: new Abstract: Agentic AI systems are increasingly being explored as production infrastructure: they reason over multiple steps, call tools, act through workflows, and

agentsarxiv-cs-ai
29 May 2026
Agents

Hijacking Agent Memory: Stealthy Trojan Attacks Through Conversational Interaction

DGX agent

arXiv:2605.29960v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly leverage long term memory to support persistent and autonomous task execution. However, this capability

agentsarxiv-cs-ai
29 May 2026
Model Releases

OpenClawBench: Benchmarking Process-side Anomalies in Real-world Agent Execution Trajectories

DGX agent

arXiv:2605.29253v1 Announce Type: new Abstract: Task success can hide process anomalies in real-world agent executions. An agent may pass the final task oracle while still accumulating unresolved ambi

model-releasesarxiv-cs-ai
29 May 2026
Agents

Towards Verifiable Multimodal Deep Research: A Multi-Agent Harness for Interleaved Report Generation

DGX agent

arXiv:2605.29861v1 Announce Type: cross Abstract: Large Language Models (LLMs) have advanced autonomous agents from deep search, which retrieves concise factual answers, to deep research, which synthe

agentsarxiv-cs-ai
29 May 2026
Model Releases

A Unified Framework for the Evaluation of LLM Agentic Capabilities

DGX agent

arXiv:2605.27898v1 Announce Type: new Abstract: As LLMs are increasingly deployed as agents, reliable assessment of their agentic capabilities has become essential. However, reported benchmark scores

model-releasesarxiv-cs-ai
28 May 2026
Agents

Adapting the Interface, Not the Model: Runtime Harness Adaptation for Deterministic LLM Agents

DGX agent

arXiv:2605.22166v2 Announce Type: replace Abstract: LLM agents are shaped not only by their language models, but also by the runtime harness that mediates observation, tool use, action execution, feed

agentsarxiv-cs-ai
28 May 2026
Model Releases

Ask Now, Use Later: Benchmarking the Proactivity Gap in Long-Lived LLM Agents

DGX agent

arXiv:2605.28108v1 Announce Type: new Abstract: A long-lived LLM agent, such as OpenClaw, earns its value by acting on a user's preferences and constraints across sessions, not just the current reques

model-releasesarxiv-cs-cl
28 May 2026
Model Releases

EgoBench: An Interactive Egocentric Multimodal Benchmark for Tool-Using Agents

DGX agent

arXiv:2605.27820v1 Announce Type: new Abstract: As AI agents increasingly operate in open, real-world environments, they require a deep synergy of multimodal perception, tool invocation with multi-hop

model-releasesarxiv-cs-ai
28 May 2026
Agents

From Instructor to Collaborator: What a 90-Participant Study Reveals about Human-Agent Collaboration in a Mobile Serious Game

DGX agent

arXiv:2605.27384v1 Announce Type: cross Abstract: This position paper reflects empirical data collected during my PhD from a large-scale within-subjects study (N = 90). The study compared a highly hum

agentsarxiv-cs-ai
28 May 2026
Agents

GUI-CIDER: Mid-training GUI Agents via Causal Internalization and Density-aware Exemplar Reselection

DGX agent

arXiv:2605.28534v1 Announce Type: new Abstract: Despite the rapid progress of multimodal large language models in building Graphical User Interface (GUI) agents, their real-world task completion is fu

agentsarxiv-cs-cl
28 May 2026
Agents

Adaptation-Free Heterogeneous Collaborative Perception with Unseen Agent Configurations

DGX agent

arXiv:2605.26642v1 Announce Type: new Abstract: Collaborative perception improves 3D object detection by enabling agents to share complementary observations, but most existing methods assume fixed or

agentsarxiv-cs-cv
27 May 2026
Model Releases

DIANOIA: Diagnostic Decomposition and Joint Optimization for Multi-Agent Reasoning

DGX agent

arXiv:2602.08586v3 Announce Type: replace Abstract: Multi-agent LLM systems consistently outperform single-agent baselines, yet practitioners still cannot predict which design works for a new task or

model-releasesarxiv-cs-ai
27 May 2026
Safety

Efficient Agentic Reinforcement Learning with On-Policy Intrinsic Knowledge Boundary Enhancement

DGX agent

arXiv:2605.26952v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) has proven effective for training LLM-based agents with external tool-use capabilities. However, we identify that ag

safetyarxiv-cs-cl
27 May 2026
Agents

EvoEmo: Towards Evolved Emotional Policies for Adversarial LLM Agents in Multi-Turn Price Negotiation

DGX agent

arXiv:2509.04310v4 Announce Type: replace Abstract: Recent research on Chain-of-Thought (CoT) reasoning in Large Language Models (LLMs) has demonstrated that agents can engage in extit{complex}, extit

agentsarxiv-cs-ai
27 May 2026
Agents

Lessons from Penetration Tests on Large-Scale Agent Systems

DGX agent

arXiv:2605.27042v1 Announce Type: cross Abstract: As AI systems gain increasing autonomy and execution capability, the number of discovered security vulnerabilities continues to rise. However, many of

agentsarxiv-cs-ai
27 May 2026
Agents

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

DGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

agentsarxiv-cs-ai
27 May 2026
Agents

QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents

DGX agent

arXiv:2605.27068v1 Announce Type: cross Abstract: Social deduction games have become a popular testbed for probing reasoning, deception, coordination, and belief modeling in Large Language Model (LLM)

agentsarxiv-cs-ai
27 May 2026
Agents

A Token/KV-Cache Communication Media Selection and Resource Allocation Strategy for Multi-Agent Collaboration

DGX agent

arXiv:2605.25422v1 Announce Type: cross Abstract: The convergence of large language models (LLMs) with 6G networks is fostering a paradigm of autonomous multi-agent cooperation, which in turn is expec

agentsarxiv-cs-ai
26 May 2026
Agents

CausalFlow: Causal Attribution and Counterfactual Repair for LLM Agent Failures

DGX agent

arXiv:2605.25338v1 Announce Type: cross Abstract: Large language model (LLM) agents frequently fail on multi-step tasks involving reasoning, tool use, and environment interaction. While such failures

agentsarxiv-cs-ai
26 May 2026
Model Releases

DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs

DGX agent

arXiv:2605.25188v1 Announce Type: new Abstract: Multi-agent LLM systems improve reasoning by combining outputs from multiple agents, but interaction-heavy methods can introduce error propagation and h

model-releasesarxiv-cs-ai
26 May 2026
Agents

HyLaT: Efficient Multi-Agent Communication via Hybrid Latent-Text Protocol

DGX agent

arXiv:2605.25421v1 Announce Type: new Abstract: Communication protocol design is a central challenge in large language model-based multi-agent systems. Existing single-channel approaches face an inher

agentsarxiv-cs-cl
26 May 2026
Agents

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

DGX agent

arXiv:2605.24775v1 Announce Type: new Abstract: Operating LLMs as coordinated multi-agent research systems over multi-hour runs surfaces failure modes that single-shot evaluation cannot: upstream prov

agentsarxiv-cs-ai
26 May 2026
Agents

Proper Scoring Rules for Agentic Uncertainty Quantification

DGX agent

arXiv:2605.24756v1 Announce Type: new Abstract: Language-model agents increasingly emit uncertainty signals throughout a trajectory, but existing agentic UQ evaluations often conflate ranking usefulne

agentsarxiv-cs-ai
26 May 2026
Safety

ARMS: Automatic Reward Shaping for Sparse-Reward Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.23562v1 Announce Type: cross Abstract: Sparse rewards are a major bottleneck in multi-agent reinforcement learning (MARL), where simultaneous learning induces non-stationarity and makes rew

safetyarxiv-cs-ai
25 May 2026
Agents

Co-ReAct: Rubrics as Step-Level Collaborators for ReAct Agents

DGX agent

arXiv:2605.23590v1 Announce Type: new Abstract: ReAct-style agents for search-intensive, multi-step reasoning tasks rely largely on their own internal judgment to decide what evidence to seek, which r

agentsarxiv-cs-ai
25 May 2026
Safety

From Correctness to Preference: A Framework for Personalized Agentic Reinforcement Learning

DGX agent

arXiv:2605.23382v1 Announce Type: new Abstract: Agentic reinforcement learning (Agentic RL) has achieved strong progress in tasks with clear success signals. However, many real-world agent application

safetyarxiv-cs-cl
25 May 2026
Model Releases

MAS-Orchestra: Understanding and Improving Multi-Agent Reasoning Through Holistic Orchestration and Controlled Benchmarks

DGX agent

arXiv:2601.14652v5 Announce Type: replace Abstract: While multi-agent systems (MAS) promise elevated intelligence through coordination of agents, current approaches to automatic MAS design under-deliv

model-releasesarxiv-cs-ai
25 May 2026
Agents

Redrawing the AI Map: A Theory of Accountability Boundaries in Agentic Ecosystems

DGX agent

arXiv:2605.23179v1 Announce Type: new Abstract: Agentic AI orchestrators reduce the interface and assembly costs of composing information systems capabilities across organizational boundaries, seeming

agentsarxiv-cs-ai
25 May 2026
Agents

TABX: A High-Throughput Sandbox Battle Simulator for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2602.01665v2 Announce Type: replace-cross Abstract: The design of environments plays a critical role in shaping the development and evaluation of cooperative multi-agent reinforcement learning (

agentsarxiv-cs-ai
25 May 2026
Model Releases

Efficient Agentic Reasoning Through Self-Regulated Simulative Planning

DGX agent

arXiv:2605.22138v1 Announce Type: cross Abstract: How should an agent decide when and how to plan? A dominant approach builds agents as reactive policies with adaptive computation (e.g., chain-of-thou

model-releasesarxiv-cs-cl
22 May 2026
Agents

From Automated to Autonomous: Hierarchical Agent-native Network Architecture (HANA)

DGX agent

arXiv:2605.20608v1 Announce Type: new Abstract: Realizing Level 4/5 Autonomous Networks (AN) demands a shift from static automation to agent-native intelligence. Current operations, reliant on rigid s

agentsarxiv-cs-ai
22 May 2026
Agents

OPERA: An Agent for Image Restoration with End-to-End Joint Planning-Execution Optimization

DGX agent

arXiv:2605.22104v1 Announce Type: new Abstract: Real-world image restoration is challenging due to complex and interacting mixed degradations. Recent agent-based approaches address this problem by com

agentsarxiv-cs-cv
22 May 2026
Agents

Auto-Dreamer: Learning Offline Memory Consolidation for Language Agents

DGX agent

arXiv:2605.20616v1 Announce Type: new Abstract: Language agents increasingly operate over streams of related tasks, yet existing memory systems struggle to convert accumulated experience into reusable

agentsarxiv-cs-cl
21 May 2026
Agents

Agentic GraphRAG: Navigating Unstructured Financial Data with Collaborative AI

DGX agent

arXiv:2605.18770v1 Announce Type: cross Abstract: We present a collaborative agentic GraphRAG framework for expert analysis of commercial registry data. Public registries are often formally accessible

agentsarxiv-cs-ai
20 May 2026
Model Releases

Conflict-Resilient Multi-Agent Reasoning via Signed Graph Modeling

DGX agent

arXiv:2605.19418v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) have demonstrated strong reasoning and decision-making capabilities that consistently surpass those of single LLM ag

model-releasesarxiv-cs-ai
20 May 2026
Agents

PAVE: A Cognitive Architecture for Legitimate Violation in Generative Agent Societies

DGX agent

arXiv:2605.19351v1 Announce Type: cross Abstract: Generative agents based on large language models reproduce believable human behavior in cooperative settings, but how they should reason in situations

agentsarxiv-cs-ai
20 May 2026
Model Releases

AgentWall: A Runtime Safety Layer for Local AI Agents

DGX agent

arXiv:2605.16265v1 Announce Type: new Abstract: The safety of autonomous AI agents is increasingly recognized as a critical open problem. As agents transition from passive text generators to active ac

model-releasesarxiv-cs-ai
19 May 2026
← Previous
1…3334353637…233
Next →