AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

TeachAnything: A Multimodal Crowdsourcing Platform for Training Embodied AI Agents in Symmetrical Reality

DGX agent

arXiv:2605.14556v1 Announce Type: new Abstract: Symmetrical Reality (SR) is emerging as a future trend for human-agent coexistence, placing higher demands on agents to acquire human-like intelligence.

agentsarxiv-cs-ai
15 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

WARD: Adversarially Robust Defense of Web Agents Against Prompt Injections

DGX agent

arXiv:2605.15030v1 Announce Type: cross Abstract: Web agents can autonomously complete online tasks by interacting with websites, but their exposure to open web environments makes them vulnerable to p

agentsarxiv-cs-ai
15 May 2026
Safety

AgenticAITA: A Proof-Of-Concept About Deliberative Multi-Agent Reasoning for Autonomous Trading Systems

DGX agent

arXiv:2605.12532v1 Announce Type: cross Abstract: Conventional algorithmic trading systems are grounded in deterministic heuristics or offline-trained statistical models that cannot adapt to the seman

safetyarxiv-cs-ai
14 May 2026
Agents

AI Harness Engineering: A Runtime Substrate for Foundation-Model Software Agents

DGX agent

arXiv:2605.13357v1 Announce Type: cross Abstract: Foundation models have transformed automated code generation, yet autonomous software-engineering agents remain unreliable in realistic development se

agentsarxiv-cs-ai
14 May 2026
Model Releases

MobiBench: Multi-Branch, Modular Benchmark for Mobile GUI Agents

DGX agent

arXiv:2512.12634v3 Announce Type: replace Abstract: Mobile GUI Agents, AI agents capable of interacting with mobile applications on behalf of users, have the potential to transform human computer inte

model-releasesarxiv-cs-ai
14 May 2026
Agents

Neurodata Without Boredom: Benchmarking Agentic AI for Data Reuse

DGX agent

arXiv:2605.12808v1 Announce Type: new Abstract: Neuroscience data are highly fragmented across labs, formats, and experimental paradigms, and reuse often requires substantial manual effort. A persiste

agentsarxiv-cs-lg
14 May 2026
Model Releases

PersonalAlign: Hierarchical Implicit Intent Alignment for Personalized GUI Agent with Long-Term User-Centric Records

DGX agent

arXiv:2601.09636v2 Announce Type: replace Abstract: While GUI agents have shown strong performance under explicit and completion instructions, real-world deployment requires aligning with users' more

model-releasesarxiv-cs-ai
14 May 2026
Agents

Predicting Decisions of AI Agents from Limited Interaction through Text-Tabular Modeling

DGX agent

arXiv:2605.12411v1 Announce Type: cross Abstract: AI agents negotiate and transact in natural language with unfamiliar counterparts: a buyer bot facing an unknown seller, or a procurement assistant ne

agentsarxiv-cs-cl
13 May 2026
Agents

SkillGen: Verified Inference-Time Agent Skill Synthesis

DGX agent

arXiv:2605.10999v1 Announce Type: new Abstract: Skills are a promising way to improve LLM agent capabilities without retraining, while keeping the added procedure reusable and controllable. However, h

agentsarxiv-cs-lg
13 May 2026
Agents

AgentSlimming: Towards Efficient and Cost-Aware Multi-Agent Systems

DGX agent

arXiv:2605.08813v1 Announce Type: new Abstract: Large Language Model-based Multi-Agent Systems (MAS) have demonstrated remarkable capabilities in complex tasks. However, manually designing optimal com

agentsarxiv-cs-lg
12 May 2026
Model Releases

Ambig-DS: A Benchmark for Task-Framing Ambiguity in Data-Science Agents

DGX agent

arXiv:2605.09698v1 Announce Type: new Abstract: As data-science agents shift from co-pilots to auto-pilots, silent misframing becomes a critical failure mode. Agents quietly commit to plausible but un

model-releasesarxiv-cs-ai
12 May 2026
Agents

Autonomous Continual Learning for Environment Adaptation of Computer-Use Agents

DGX agent

arXiv:2602.10356v2 Announce Type: replace Abstract: Real-world digital environments are highly diverse and dynamic. These characteristics cause agents to frequently encounter unseen environments and d

agentsarxiv-cs-cl
12 May 2026
Agents

Decentralized Contingency MPC based on Safe Sets for Nonlinear Multi-agent Collision Avoidance

DGX agent

arXiv:2605.10738v1 Announce Type: cross Abstract: Decentralized collision avoidance remains challenging, particularly when agents do not communicate any information related to planned trajectories. Mo

agentsarxiv-cs-ro
12 May 2026
Agents

DeepRefine: Agent-Compiled Knowledge Refinement via Reinforcement Learning

DGX agent

arXiv:2605.10488v1 Announce Type: cross Abstract: Agent-compiled knowledge bases provide persistent external knowledge for large language model (LLM) agents in open-ended, knowledge-intensive downstre

agentsarxiv-cs-ai
12 May 2026
Agents

EquiMem: Calibrating Shared Memory in Multi-Agent Debate via Game-Theoretic Equilibrium

DGX agent

arXiv:2605.09278v1 Announce Type: new Abstract: Multi-agent debate (MAD) systems increasingly rely on shared memory to support long-horizon reasoning, but this convenience opens a critical vulnerabili

agentsarxiv-cs-ai
12 May 2026
Local Ai

Fully Decentralized Cooperative Multi-Agent Reinforcement Learning is A Context Modeling Problem

DGX agent

arXiv:2509.15519v2 Announce Type: replace Abstract: This paper studies fully decentralized cooperative multi-agent reinforcement learning, where each agent solely observes the states, its local action

local-aiarxiv-cs-lg
12 May 2026
Model Releases

LITMUS: Benchmarking Behavioral Jailbreaks of LLM Agents in Real OS Environments

DGX agent

arXiv:2605.10779v1 Announce Type: cross Abstract: The rapid proliferation of LLM-based autonomous agents in real operating system environments introduces a new category of safety risk beyond content s

model-releasesarxiv-cs-cl
12 May 2026
Safety

Research on Security Enhancement Methods for Adversarial Robust Large Language Model Intelligent Agents for Medical Decision-Making Tasks

DGX agent

arXiv:2605.08257v1 Announce Type: cross Abstract: Motivated by the challenge to improve the adversarial robustness, security, and trust of medical decision making intelligent agents, this study develo

safetyarxiv-cs-ai
12 May 2026
Model Releases

Strategic Exploitation in LLM Agent Markets: A Simulation Framework for E-Commerce Trust

DGX agent

arXiv:2605.10059v1 Announce Type: new Abstract: Agent-based modeling (ABM) has long been used in economics to study human behavior, and large language model (LLM) agents now enable new forms of social

model-releasesarxiv-cs-ai
12 May 2026
Agents

A Self-Healing Framework for Reliable LLM-Based Autonomous Agents

DGX agent

arXiv:2605.06737v1 Announce Type: cross Abstract: Autonomous agents based on Large Language Models (LLMs) are increasingly being utilized in complex software systems. However, reliability remains a si

agentsarxiv-cs-ai
11 May 2026
Agents

AgentProg: Empowering Long-Horizon GUI Agents with Program-Guided Context Management

DGX agent

arXiv:2512.10371v2 Announce Type: replace Abstract: The rapid development of mobile GUI agents has stimulated growing research interest in long-horizon task automation. However, building agents for th

agentsarxiv-cs-ai
11 May 2026
Agents

Belief Memory: Agent Memory Under Partial Observability

DGX agent

arXiv:2605.05583v2 Announce Type: replace Abstract: LLM agents that operate over long context depend on external memory to accumulate knowledge over time. However, existing methods typically store eac

agentsarxiv-cs-ai
11 May 2026
Model Releases

Exact Is Easier: Credit Assignment for Cooperative LLM Agents

DGX agent

arXiv:2603.06859v2 Announce Type: replace-cross Abstract: Removing an agent from a cooperative team to measure its contribution seems natural, yet in multi-agent LLM systems this evaluation distorts t

model-releasesarxiv-cs-ai
11 May 2026
Agents

Group of Skills: Group-Structured Skill Retrieval for Agent Skill Libraries

DGX agent

arXiv:2605.06978v1 Announce Type: cross Abstract: Skill-augmented agents increasingly rely on large reusable skill libraries, but retrieving relevant skills is not the same as presenting usable contex

agentsarxiv-cs-ai
11 May 2026
Local Ai

Learning to Communicate Locally for Large-Scale Multi-Agent Pathfinding

DGX agent

arXiv:2605.07637v1 Announce Type: new Abstract: Multi-agent pathfinding (MAPF) is a widely used abstraction for multi-robot trajectory planning problems, where multiple homogeneous agents move simulta

local-aiarxiv-cs-ai
11 May 2026
Model Releases

SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks

DGX agent

arXiv:2603.24755v2 Announce Type: replace-cross Abstract: Software development is iterative, yet agentic coding benchmarks hide design issues through their single-shot setup. Recent iterative benchmar

model-releasesarxiv-cs-ai
11 May 2026
Agents

TraceFix: Repairing Agent Coordination Protocols with TLA+ Counterexamples

DGX agent

arXiv:2605.07935v1 Announce Type: new Abstract: We present TraceFix, a verification-first pipeline for Large Language Model (LLM) multi-agent coordination. An agent synthesizes a protocol topology as

agentsarxiv-cs-ai
11 May 2026
Agents

When Stored Evidence Stops Being Usable: Scale-Conditioned Evaluation of Agent Memory

DGX agent

arXiv:2605.07313v1 Announce Type: new Abstract: Memory-agent evaluations report fixed-snapshot accuracy or retrieval quality, but these scores do not show whether evidence remains usable as irrelevant

agentsarxiv-cs-ai
11 May 2026
Safety

Why Does Agentic Safety Fail to Generalize Across Tasks?

DGX agent

arXiv:2605.06992v1 Announce Type: new Abstract: AI agents are increasingly deployed in multi-task settings, where the task to perform is specified at test time, and the agent must generalize to unseen

safetyarxiv-cs-lg
11 May 2026
Agents

Pact: A Choreographic Language for Agentic Ecosystems

DGX agent

arXiv:2605.03143v1 Announce Type: cross Abstract: Recent advances in large language models have led to the rise of software systems (i.e. agents) that execute with increasing autonomy on behalf of use

agentsarxiv-cs-ai
7 May 2026
Agents

HeavySkill: Heavy Thinking as the Inner Skill in Agentic Harness

DGX agent

arXiv:2605.02396v1 Announce Type: new Abstract: Recent advances in agentic harness with orchestration frameworks that coordinate multiple agents with memory, skills, and tool use have achieved remarka

agentsarxiv-cs-ai
6 May 2026
Agents

Hybrid Inspection and Task-Based Access Control in Zero-Trust Agentic AI

DGX agent

arXiv:2605.02682v1 Announce Type: new Abstract: Authorizing Large Language Model (LLM)-driven agents to dynamically invoke tools and access protected resources introduces significant security risks, a

agentsarxiv-cs-ai
6 May 2026
Agents

Less Interaction But More Explanation: A Communication Perspective on Agentic AI Interfaces

DGX agent

arXiv:2605.01610v1 Announce Type: cross Abstract: AI systems have long been expected to interact with users, answering questions, generating content, and continuing (social) conversations. Agentic AI,

agentsarxiv-cs-ai
6 May 2026
Agents

Lifting Traces to Logic: Programmatic Skill Induction with Neuro-Symbolic Learning for Long-Horizon Agentic Tasks

DGX agent

arXiv:2605.01293v1 Announce Type: new Abstract: Foundation model-driven agents often struggle with long-horizon planning due to the transient nature of purely prompting-based reasoning. While existing

agentsarxiv-cs-ai
6 May 2026
Agents

OpenSeeker-v2: Pushing the Limits of Search Agents with Informative and High-Difficulty Trajectories

DGX agent

arXiv:2605.04036v1 Announce Type: cross Abstract: Deep search capabilities have become an indispensable competency for frontier Large Language Model (LLM) agents, yet their development remains dominat

agentsarxiv-cs-cl
6 May 2026
Safety

Ambient Persuasion in a Deployed AI Agent: Unauthorized Escalation Following Routine Non-Adversarial Content Exposure

DGX agent

arXiv:2605.00055v1 Announce Type: cross Abstract: We report a safety incident in a deployed multi-agent research system in which a primary AI agent installed 107 unauthorized software components, over

safetyarxiv-cs-ai
5 May 2026
Agents

Optimized and kinematically feasible multi-agent motion planning

DGX agent

arXiv:2605.01996v1 Announce Type: new Abstract: Multi-agent motion planning (MAMP) is an important problem for autonomous systems with multiple agents. In this work we propose a two-step method for fi

agentsarxiv-cs-ro
5 May 2026
Model Releases

SciResearcher: Scaling Deep Research Agents for Frontier Scientific Reasoning

DGX agent

arXiv:2605.01489v1 Announce Type: cross Abstract: Frontier scientific reasoning is rapidly emerging as a key foundation for advancing AI agents in automated scientific discovery. Deep research agents

model-releasesarxiv-cs-cl
5 May 2026
Agents

Position: agentic AI orchestration should be Bayes-consistent

DGX agent

arXiv:2605.00742v1 Announce Type: cross Abstract: LLMs excel at predictive tasks and complex reasoning tasks, but many high-value deployments rely on decisions under uncertainty, for example, which to

agentsarxiv-cs-lg
4 May 2026
Model Releases

Exploring Interaction Paradigms for LLM Agents in Scientific Visualization

DGX agent

arXiv:2604.27996v1 Announce Type: new Abstract: This paper examines how different types of large language model (LLM) agents perform on scientific visualization (SciVis) tasks, where users generate vi

model-releasesarxiv-cs-ai
1 May 2026
Agents

Modeling Clinical Concern Trajectories in Language Model Agents

DGX agent

arXiv:2604.27872v1 Announce Type: new Abstract: Large language model (LLM) agents deployed in clinical settings often exhibit abrupt, threshold-driven behavior, offering little visibility into accumul

agentsarxiv-cs-ai
1 May 2026
Agents

Rethinking Agentic Reinforcement Learning In Large Language Models

DGX agent

arXiv:2604.27859v1 Announce Type: new Abstract: Reinforcement Learning (RL) has traditionally focused on training specialized agents to optimize predefined reward functions within narrowly defined env

agentsarxiv-cs-ai
1 May 2026
Safety

Safe Bilevel Delegation (SBD): A Formal Framework for Runtime Delegation Safety in Multi-Agent Systems

DGX agent

arXiv:2604.27358v1 Announce Type: new Abstract: As large language model (LLM) agents are deployed in high-stakes environments, the question of how safely to delegate subtasks to specialized sub-agents

safetyarxiv-cs-ai
1 May 2026
Agents

Security Attack and Defense Strategies for Autonomous Agent Frameworks: A Layered Review with OpenClaw as a Case Study

DGX agent

arXiv:2604.27464v1 Announce Type: cross Abstract: Autonomous agent frameworks built upon large language models (LLMs) are evolving into complex, tool-integrated, and continuously operating systems, in

agentsarxiv-cs-ai
1 May 2026
Agents

Self-Evolving Software Agents

DGX agent

arXiv:2604.27264v1 Announce Type: cross Abstract: Autonomous agents can adapt their behaviour to changing environments, but remain bound to requirements, goals, and capabilities fixed at design time,

agentsarxiv-cs-ai
1 May 2026
Agents

A Survey of Multi-Agent Deep Reinforcement Learning with Graph Neural Network-Based Communication

DGX agent

arXiv:2604.25972v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), the integration of a communication mechanism, allowing agents to better learn to coordinate their action

agentsarxiv-cs-ai
30 Apr 2026
Agents

GLM-5V-Turbo: Toward a Native Foundation Model for Multimodal Agents

DGX agent

arXiv:2604.26752v1 Announce Type: new Abstract: We present GLM-5V-Turbo, a step toward native foundation models for multimodal agents. As foundation models are increasingly deployed in real environmen

agentsarxiv-cs-cv
30 Apr 2026
Model Releases

LATTICE: Evaluating Decision Support Utility of Crypto Agents

DGX agent

arXiv:2604.26235v1 Announce Type: cross Abstract: We introduce LATTICE, a benchmark for evaluating the decision support utility of crypto agents in realistic user-facing scenarios. Prior crypto agent

model-releasesarxiv-cs-ai
30 Apr 2026
← Previous
1…2829303132…233
Next →