AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

COOP^2: Defining, Observing, and Repairing Cooperation in LLM Multi-Agent Systems

DGX agent

arXiv:2603.00349v2 Announce Type: replace Abstract: Many complex tasks require extended effort, diverse capabilities, or coordinated actions beyond what a single agent can provide. However, simply add

agentsarxiv-cs-ai
28 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Learn from Weaknesses: Automated Domain Specialization for Small Computer-Use Agents

DGX agent

arXiv:2605.28775v1 Announce Type: cross Abstract: Computer-use agents (CUAs) have recently made substantial progress, but deploying a separate large expert for each software domain remains expensive.

agentsarxiv-cs-ai
28 May 2026
Safety

Mobile-Aptus: Confidence-Driven Proactive and Robust Interaction in MLLM-based Mobile-Using Agents

DGX agent

arXiv:2605.28629v1 Announce Type: new Abstract: Recent advancements in multimodal large language models (MLLMs) have shown exceptional potential in enabling mobile-using agents to autonomously execute

safetyarxiv-cs-cl
28 May 2026
Agents

Out of Sight, Not Out of Mind: Unveiling Latent Attack in Latent-based Multi-Agent Systems

DGX agent

arXiv:2605.28214v1 Announce Type: cross Abstract: Latent-based multi-agent systems replace parts of explicit inter-agent communication with hidden representations, offering a new direction for efficie

agentsarxiv-cs-lg
28 May 2026
Agents

Skill-as-Pseudocode: Refactoring Skill Libraries to Pseudocode for LLM Agents

DGX agent

arXiv:2605.27955v1 Announce Type: cross Abstract: Markdown skill libraries for LLM agents ship as free-form prose, forcing the agent to re-derive both the input schema and the concrete invocation synt

agentsarxiv-cs-cl
28 May 2026
Agents

Governed Evolution of Agent Runtimes through Executable Operational Cognition

DGX agent

arXiv:2605.27328v1 Announce Type: cross Abstract: Recent advances in agentic systems increasingly treat code as an executable operational substrate rather than as a disposable output artifact. Prior w

agentsarxiv-cs-ai
27 May 2026
Agents

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation

DGX agent

arXiv:2605.27366v1 Announce Type: new Abstract: Large language model (LLM) agents rely on reusable skills to solve complex tasks. However, existing skill creation approaches treat skills as isolated a

agentsarxiv-cs-ai
27 May 2026
Model Releases

Persona2Web: Benchmarking Personalized Web Agents for Contextual Reasoning with User History

DGX agent

arXiv:2602.17003v2 Announce Type: replace-cross Abstract: Large language models have advanced web agents, yet current agents lack personalization capabilities. Since users rarely specify every detail

model-releasesarxiv-cs-ai
27 May 2026
Safety

Securing Multi-Agent Systems Against Corruptions via Node Contribution Backpropagation

DGX agent

arXiv:2510.19420v2 Announce Type: replace-cross Abstract: Multi-Agent Systems (MAS) have become a prevalent paradigm for Large Language Model (LLM) applications. However, the complex multi-agent desig

safetyarxiv-cs-ai
27 May 2026
Agents

DemoEvolve: Overcoming Sparse Feedback in Agentic Harness Evolution with Demonstrations

DGX agent

arXiv:2605.24539v1 Announce Type: new Abstract: Agent harness evolution improves frozen language-model agents by modifying the executable structures around them. We study this paradigm as a form of sa

agentsarxiv-cs-ai
26 May 2026
Agents

SPARK: Search Personalization via Agent-Driven Retrieval and Knowledge-sharing

DGX agent

arXiv:2512.24008v3 Announce Type: replace Abstract: Personalized search demands the ability to model users' evolving, multi-dimensional information needs; a challenge for systems constrained by static

agentsarxiv-cs-ai
26 May 2026
Agents

MemAudit: Post-hoc Auditing of Poisoned Agent Memory via Causal Attribution and Structural Anomaly Detection

DGX agent

arXiv:2605.23723v1 Announce Type: new Abstract: Large language model agents increasingly rely on persistent memory to store past interactions, retrieve relevant demonstrations, and improve long-horizo

agentsarxiv-cs-ai
25 May 2026
Agents

When Planning Fails Despite Correct Execution: On Epistemic Calibration for LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.23414v1 Announce Type: new Abstract: LLM-based multi-agent systems can fail even when planned actions are executed correctly because agents may misjudge their knowledge when evaluating plan

agentsarxiv-cs-ai
25 May 2026
Agents

Mix-Quant: Quantized Prefilling, Precise Decoding for Agentic LLMs

DGX agent

arXiv:2605.20315v1 Announce Type: new Abstract: LLM agents have recently emerged as a powerful paradigm for solving complex tasks through planning, tool use, memory retrieval, and multi-step interacti

agentsarxiv-cs-cl
21 May 2026
Agents

When Skills Don't Help: A Negative Result on Procedural Knowledge for Tool-Grounded Agents in Offensive Cybersecurity

DGX agent

arXiv:2605.20023v1 Announce Type: new Abstract: Agent Skills, structured packages of procedural knowledge loaded into an LLM agent at inference time, are widely reported to improve task pass rates by

agentsarxiv-cs-ai
20 May 2026
Agents

Can LLM Agents Be CFOs? Benchmarking Long-Horizon Resource Allocation in an Uncertain Enterprise Environment

DGX agent

arXiv:2603.23638v2 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly tested on complex tasks, but their ability to allocate scarce resources over long horizons remain

agentsarxiv-cs-ai
19 May 2026
Agents

GRASP: Graph Agentic Search over Propositions for Multi-hop Question Answering

DGX agent

arXiv:2605.16598v1 Announce Type: cross Abstract: Agentic retrieval improves multi-hop question answering by giving language models autonomy to iteratively gather evidence. Recent work augments these

agentsarxiv-cs-ai
19 May 2026
Safety

Helpful to a Fault: Measuring Illicit Assistance in Multi-Turn, Multilingual LLM Agents

DGX agent

arXiv:2602.16346v3 Announce Type: replace Abstract: LLM-based agents execute real-world workflows via tools and memory. These affordances enable ill-intended adversaries to also use these agents to ca

safetyarxiv-cs-cl
19 May 2026
Model Releases

MetaCogAgent: A Metacognitive Multi-Agent LLM Framework with Self-Aware Task Delegation

DGX agent

arXiv:2605.17292v1 Announce Type: new Abstract: Multi-agent large language model (LLM) systems have shown promise for solving complex tasks through agent collaboration. However, existing frameworks as

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

NeuroMAS: Multi-Agent Systems as Neural Networks with Joint Reinforcement Learning

DGX agent

arXiv:2605.16757v1 Announce Type: new Abstract: Multi-agent language systems are often built as hand-designed workflows, where agents are assigned semantic roles and communication protocols are specif

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Visual Agentic Memory: Enabling Online Long Video Understanding via Online Indexing, Hierarchical Memory, and Agentic Retrieval

DGX agent

arXiv:2605.16481v1 Announce Type: cross Abstract: Long video understanding requires more than large context windows. It also needs a memory mechanism that decides what visual evidence to retain, keeps

model-releasesarxiv-cs-ai
19 May 2026
Agents

Differentiable Mixture-of-Agents Incentivizes Swarm Intelligence of Large Language Models

DGX agent

arXiv:2605.15706v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have catalyzed the development of multi-agent systems (MAS) for complex reasoning tasks. However, existi

agentsarxiv-cs-lg
18 May 2026
Agents

H-Mem: A Novel Memory Mechanism for Evolving and Retrieving Agent Memory via a Hybrid Structure

DGX agent

arXiv:2605.15701v1 Announce Type: cross Abstract: Memory data are ubiquitous in Large Language Model (LLM)-based agents (e.g., OpenClaw and Manus). A few recent works have attempted to exploit agents'

agentsarxiv-cs-ai
18 May 2026
Agents

Look Before You Leap: Autonomous Exploration for LLM Agents

DGX agent

arXiv:2605.16143v1 Announce Type: new Abstract: Large language model based agents often fail in unfamiliar environments due to premature exploitation: a tendency to act on prior knowledge before acqui

agentsarxiv-cs-ai
18 May 2026
Agents

Known By Their Actions: Fingerprinting LLM Browser Agents via UI Traces

DGX agent

arXiv:2605.14786v1 Announce Type: cross Abstract: As LLM-based agents increasingly browse the web on users' behalf, a natural question arises: can websites passively identify which underlying model po

agentsarxiv-cs-ai
15 May 2026
Model Releases

Web Agents Should Adopt the Plan-Then-Execute Paradigm

DGX agent

arXiv:2605.14290v1 Announce Type: cross Abstract: ReAct has become the default architecture across LLM agents, and many existing web agents follow this paradigm. We argue that it is the wrong default

model-releasesarxiv-cs-ai
15 May 2026
Agents

AgenticFlict: A Large-Scale Dataset of Merge Conflicts in AI Coding Agent Pull Requests on GitHub

DGX agent

arXiv:2604.03551v2 Announce Type: replace-cross Abstract: Software Engineering 3.0 marks a paradigm shift in software development, in which AI coding agents are no longer just assistive tools but acti

agentsarxiv-cs-ai
14 May 2026
Agents

CHAL: Council of Hierarchical Agentic Language

DGX agent

arXiv:2605.12718v1 Announce Type: new Abstract: Multi-agent debate has emerged as a promising approach for improving LLM reasoning on ground-truth tasks, yet current methodologies face certain structu

agentsarxiv-cs-ai
14 May 2026
Model Releases

Embodied Multi-Agent Coordination by Aligning World Models Through Dialogue

DGX agent

arXiv:2605.12920v1 Announce Type: cross Abstract: Effective collaboration between embodied agents requires more than acting in a shared environment; it demands communication grounded in each agent's e

model-releasesarxiv-cs-ai
14 May 2026
Agents

OpenAaaS: An Open Agent-as-a-Service Framework for Distributed Materials-Informatics Research

DGX agent

arXiv:2605.13618v1 Announce Type: cross Abstract: The Materials Genome Initiative catalyzed the proliferation of centralized platforms--SaaS, PaaS, and IaaS--that aggregate computational and experimen

agentsarxiv-cs-ai
14 May 2026
Safety

The Horizon Threshold in Cooperative Multi-Agent Reward-Free Exploration

DGX agent

arXiv:2602.01453v3 Announce Type: replace Abstract: We study cooperative multi-agent reinforcement learning in the setting of reward-free exploration, where multiple agents jointly explore an unknown

safetyarxiv-cs-lg
14 May 2026
Agents

Focusing Influence Mechanism for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2506.19417v2 Announce Type: replace Abstract: Cooperative multi-agent reinforcement learning (MARL) under sparse rewards remains fundamentally challenging because agents often fail to concentrat

agentsarxiv-cs-lg
13 May 2026
Safety

EvoMAS: Learning Execution-Time Workflows for Multi-Agent Systems

DGX agent

arXiv:2605.08769v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential on complex tasks through agent specialization, tool use, and collaborat

safetyarxiv-cs-ai
12 May 2026
Model Releases

GAMBIT: A Three-Mode Benchmark for Adversarial Robustness in Multi-Agent LLM Collectives

DGX agent

arXiv:2605.09027v1 Announce Type: new Abstract: In multi-agent systems (MAS), a single deceptive agent can nullify all gains of an agentic AI collective and evade deployed defenses. However, existing

model-releasesarxiv-cs-cl
12 May 2026
Agents

MIND-Skill: Quality-Guaranteed Skill Generation via Multi-Agent Induction and Deduction

DGX agent

arXiv:2605.08670v1 Announce Type: new Abstract: Large language model (LLM) powered AI agents have emerged as a promising paradigm for autonomous problem-solving, yet they continue to struggle with com

agentsarxiv-cs-ai
12 May 2026
Agents

Robust Multi-Agent LLMs under Byzantine Faults

DGX agent

arXiv:2605.09076v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly collaborate over peer-to-peer networks to improve their reliability. However, these same interactions c

agentsarxiv-cs-ai
12 May 2026
Agents

WebTrap: Stealthy Mid-Task Hijacking of Browser Agents During Navigation

DGX agent

arXiv:2605.08310v1 Announce Type: cross Abstract: Browser agents are increasingly deployed in long-horizon tasks, which require executing extended action chains to accomplish user goals. However, this

agentsarxiv-cs-ai
12 May 2026
Agents

Willful Disobedience: Automatically Detecting Failures in Agentic Traces

DGX agent

arXiv:2603.23806v2 Announce Type: replace-cross Abstract: AI agents are increasingly embedded in real software systems, where they execute multi-step workflows through multi-turn dialogue, tool invoca

agentsarxiv-cs-ai
12 May 2026
Agents

SOM: Structured Opponent Modeling for LLM-based Agents via Structural Causal Model

DGX agent

arXiv:2605.07301v1 Announce Type: new Abstract: Accurately predicting opponents' behavior from interactions is a fundamental capability for large language model (LLM)-based agents in multi-agent and g

agentsarxiv-cs-ai
11 May 2026
Agents

WebClipper: Efficient Evolution of Web Agents with Graph-based Trajectory Pruning

DGX agent

arXiv:2602.12852v2 Announce Type: replace Abstract: Deep Research systems based on web agents have shown strong potential in solving complex information-seeking tasks, yet their search efficiency rema

agentsarxiv-cs-ai
11 May 2026
Agents

Agentic publications: redesigning scientific publishing in the age of thinking large language models

DGX agent

arXiv:2505.13246v2 Announce Type: replace Abstract: Purpose: This paper introduces the concept of 'Agentic Publication,' a novel LLM-driven framework designed to complement traditional scientific publ

agentsarxiv-cs-ai
7 May 2026
Agents

OpenSearch-VL: An Open Recipe for Frontier Multimodal Search Agents

DGX agent

arXiv:2605.05185v1 Announce Type: new Abstract: Deep search has become a crucial capability for frontier multimodal agents, enabling models to solve complex questions through active search, evidence v

agentsarxiv-cs-cv
7 May 2026
Agents

A Low-Latency Fraud Detection Layer for Detecting Adversarial Interaction Patterns in LLM-Powered Agents

DGX agent

arXiv:2605.01143v1 Announce Type: new Abstract: Large Language Model (LLM)-powered agents demonstrate strong capabilities in autonomous task execution, tool use, and multi-step reasoning. However, the

agentsarxiv-cs-ai
6 May 2026
Agents

Beyond State Machines: Executing Network Procedures with Agentic Tool-Calling Sequences

DGX agent

arXiv:2605.02584v1 Announce Type: cross Abstract: Agentic AI will be an essential enabling technology for designing future mobile communication systems, which could provide flexible and customized ser

agentsarxiv-cs-ai
6 May 2026
Safety

Catching the Infection Before It Spreads: Foresight-Guided Defense in Multi-Agent Systems

DGX agent

arXiv:2605.01758v1 Announce Type: new Abstract: Large multimodal model-based Multi-Agent Systems (MASs) enable collaborative complex problem solving through specialized agents. However, MASs are vulne

safetyarxiv-cs-ai
6 May 2026
Agents

CoFlow: Coordinated Few-Step Flow for Offline Multi-Agent Decision Making

DGX agent

arXiv:2605.01457v1 Announce Type: new Abstract: Generative models have emerged as a major paradigm for offline multi-agent reinforcement learning (MARL), but existing approaches require many iterative

agentsarxiv-cs-ai
6 May 2026
Agents

Foundation-Model-Based Agents in Industrial Automation: Purposes, Capabilities, and Open Challenges

DGX agent

arXiv:2605.02592v1 Announce Type: new Abstract: Foundation models, particularly large language models, are increasingly integrated into agent architectures for industrial tasks such as decision suppor

agentsarxiv-cs-ai
6 May 2026
Model Releases

Towards Multi-Agent Autonomous Reasoning in Hydrodynamics

DGX agent

arXiv:2605.01102v1 Announce Type: new Abstract: Single-agent systems (SAS) have become the default pattern for LLM-driven scientific workflows, but routing planning, tool use, and synthesis through a

model-releasesarxiv-cs-ai
6 May 2026
← Previous
1…2223242526…233
Next →