AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

1GC-7RC: One Graphic Card -- Seven Research Challenges! How Good Are AI Agents at Doing Your Job?

DGX agent

arXiv:2605.17046v1 Announce Type: cross Abstract: Autonomous AI coding agents are becoming a core tool for ML practitioners in industry and research alike. Despite this growing adoption, no standardiz

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

ASPI: Seeking Ambiguity Clarification Amplifies Prompt Injection Vulnerability in LLM Agents

DGX agent

arXiv:2605.17324v1 Announce Type: cross Abstract: Clarification-seeking behavior is widely regarded as a desirable property of LLM agents, enabling them to resolve ambiguity before acting on underspec

model-releasesarxiv-cs-ai
19 May 2026
Agents

From Prompts to Protocols: An AI Agent for Laboratory Automation

DGX agent

arXiv:2605.16552v1 Announce Type: new Abstract: Automating science laboratories enables faster, safer, more accurate, and more reproducible execution of protocols, accelerating the discovery and testi

agentsarxiv-cs-ai
19 May 2026
Agents

Heterogeneous Information-Bottleneck Coordination Graphs for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.17393v1 Announce Type: new Abstract: Coordination graphs are a central abstraction in cooperative multi-agent reinforcement learning (MARL), yet existing sparse-graph learners lack a theore

agentsarxiv-cs-ai
19 May 2026
Agents

Lying with Truths: Open-Channel Multi-Agent Collusion for Belief Manipulation via Generative Montage

DGX agent

arXiv:2601.01685v2 Announce Type: replace-cross Abstract: As large language models (LLMs) transition to autonomous agents synthesizing real-time information, their reasoning capabilities introduce an

agentsarxiv-cs-ai
19 May 2026
Model Releases

Overeager Coding Agents: Measuring Out-of-Scope Actions on Benign Tasks

DGX agent

arXiv:2605.18583v1 Announce Type: cross Abstract: Coding agents now run autonomously with shell, file, and network privileges. When a user issues a benign request, the agent sometimes does more than a

model-releasesarxiv-cs-ai
19 May 2026
Agents

SEMA-RAG: A Self-Evolving Multi-Agent Retrieval-Augmented Generation Framework for Medical Reasoning

DGX agent

arXiv:2605.17101v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) is widely employed to mitigate risks such as hallucinations and knowledge obsolescence in medical question answer

agentsarxiv-cs-ai
19 May 2026
Agents

The Alpha Illusion: Reported Alpha from LLM Trading Agents Should Not Be Treated as Deployment Evidence

DGX agent

arXiv:2605.16895v1 Announce Type: cross Abstract: End-to-end LLM trading agents have moved quickly from research curiosity to a small ecosystem of named systems, including FinCon, FinMem, TradingAgent

agentsarxiv-cs-ai
19 May 2026
Local Ai

AgentStop: Terminating Local AI Agents Early to Save Energy in Consumer Devices

DGX agent

arXiv:2605.15206v1 Announce Type: cross Abstract: Autonomous agents powered by large language models (LLMs) are increasingly used to automate complex, multi-step tasks such as coding or web-based ques

local-aiarxiv-cs-ai
18 May 2026
Local Ai

Is Agentic AI Ready for Real-World Hardware Engineering? A Deep Dive with Phoenix-bench

DGX agent

arXiv:2605.15226v1 Announce Type: cross Abstract: We ask whether agentic AI systems built for software engineering transfer to realistic hardware engineering. Existing hardware LLM benchmarks isolate

local-aiarxiv-cs-ai
18 May 2026
Safety

SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization

DGX agent

arXiv:2604.02268v2 Announce Type: replace Abstract: Agent skills, structured packages of procedural knowledge and executable resources that agents dynamically load at inference time, have become a rel

safetyarxiv-cs-lg
18 May 2026
Agents

Traj-CoA: Patient Trajectory Modeling via Chain-of-Agents for Lung Cancer Risk Prediction

DGX agent

arXiv:2510.10454v2 Announce Type: replace Abstract: Large language models (LLMs) offer a generalizable approach for modeling patient trajectories, but suffer from the long and noisy nature of electron

agentsarxiv-cs-ai
18 May 2026
Agents

A Two-Dimensional Framework for AI Agent Design Patterns: Cognitive Function and Execution Topology

DGX agent

arXiv:2605.13850v1 Announce Type: new Abstract: Existing frameworks for LLM-based agent architectures describe systems from a single perspective: industry guides (Anthropic, Google, LangChain) focus o

agentsarxiv-cs-ai
15 May 2026
Agents

ChromaFlow: A Negative Ablation Study of Orchestration Overhead in Tool-Augmented Agent Evaluation

DGX agent

arXiv:2605.14102v1 Announce Type: new Abstract: Autonomous language-model agents increasingly combine planning, tool use, document processing, browsing, code execution, and verification loops. These c

agentsarxiv-cs-ai
15 May 2026
Model Releases

Collider-Bench: Benchmarking AI Agents with Particle Physics Analysis Reproduction

DGX agent

arXiv:2605.13950v1 Announce Type: cross Abstract: Autonomous language-model agents are increasingly evaluated on long-horizon tool-use tasks, but existing benchmarks rarely capture the complexity and

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Invisible Orchestrators Suppress Protective Behavior and Dissociate Power-Holders: Safety Risks in Multi-Agent LLM Systems

DGX agent

arXiv:2605.13851v1 Announce Type: new Abstract: Multi-agent orchestration -- in which a hidden coordinator manages specialized worker agents -- is becoming the default architecture for enterprise AI d

model-releasesarxiv-cs-ai
15 May 2026
Agents

MALLVI: A Multi-Agent Framework for Integrated Generalized Robotics Manipulation

DGX agent

arXiv:2602.16898v5 Announce Type: replace-cross Abstract: Task planning for robotic manipulation with large language models (LLMs) is an emerging area. Prior approaches rely on specialized models, fin

agentsarxiv-cs-ai
15 May 2026
Model Releases

Near-Miss: Latent Policy Failure Detection in Agentic Workflows

DGX agent

arXiv:2603.29665v2 Announce Type: replace Abstract: Agentic systems for business process automation often require compliance with policies governing conditional updates to the system state. Evaluation

model-releasesarxiv-cs-cl
15 May 2026
Agents

PREPING: Building Agent Memory without Tasks

DGX agent

arXiv:2605.13880v1 Announce Type: new Abstract: Agent memory is typically constructed either offline from curated demonstrations or online from post-deployment interactions. However, regardless of how

agentsarxiv-cs-ai
15 May 2026
Agents

Ready from Day 1: Population-Aware Coordination for Large-Scale Constrained Multi-Agent Systems

DGX agent

arXiv:2605.13900v1 Announce Type: cross Abstract: In large-scale multi-agent systems with shared resource constraints, an upstream planner must iteratively evaluate candidate resource plans -- assessi

agentsarxiv-cs-lg
15 May 2026
Safety

SimPersona: Learning Discrete Buyer Personas from Raw Clickstreams for Grounded E-Commerce Agents

DGX agent

arXiv:2605.14205v1 Announce Type: new Abstract: LLM-based web agents can navigate live storefronts, yet they often collapse to a single 'average buyer' policy, failing to capture the heterogeneous and

safetyarxiv-cs-ai
15 May 2026
Model Releases

SWE-Chain: Benchmarking Coding Agents on Chained Release-Level Package Upgrades

DGX agent

arXiv:2605.14415v1 Announce Type: cross Abstract: Coding agents powered by large language models are increasingly expected to perform realistic software maintenance tasks beyond isolated issue resolut

model-releasesarxiv-cs-ai
15 May 2026
Agents

Video2GUI: Synthesizing Large-Scale Interaction Trajectories for Generalized GUI Agent Pretraining

DGX agent

arXiv:2605.14747v1 Announce Type: cross Abstract: Recent advances in multimodal large language models have driven growing interest in graphical user interface (GUI) agents, yet their generalization re

agentsarxiv-cs-ai
15 May 2026
Agents

Why Neighborhoods Matter: Traversal Context and Provenance in Agentic GraphRAG

DGX agent

arXiv:2605.15109v1 Announce Type: new Abstract: Retrieval-Augmented Generation can improve factuality by grounding answers in external evidence, but Agentic GraphRAG complicates what it means for cita

agentsarxiv-cs-ai
15 May 2026
Model Releases

Building Interactive Real-Time Agents with Asynchronous I/O and Speculative Tool Calling

DGX agent

arXiv:2605.13360v1 Announce Type: new Abstract: There is a growing demand for agentic AI technologies for a range of downstream applications like customer service and personal assistants. For applicat

model-releasesarxiv-cs-lg
14 May 2026
Agents

IdeaForge: A Knowledge Graph-Grounded Multi-Agent Framework for Cross-Methodology Innovation Analysis and Patent Claim Generation

DGX agent

arXiv:2605.13311v1 Announce Type: new Abstract: Current AI-assisted innovation systems typically apply a single ideation methodology (such as TRIZ or Design Thinking) using sequential prompt-based wor

agentsarxiv-cs-ai
14 May 2026
Agents

Position: Agentic AI System Is a Foreseeable Pathway to AGI

DGX agent

arXiv:2605.12966v1 Announce Type: new Abstract: Is monolithic scaling the only path to AGI? This paper challenges the dogma that purely scaling a single model is sufficient to achieve Artificial Gener

agentsarxiv-cs-ai
14 May 2026
Agents

AgentShield: Deception-based Compromise Detection for Tool-using LLM Agents

DGX agent

arXiv:2605.11026v1 Announce Type: cross Abstract: Defenses against indirect prompt injection (IPI) in tool-using LLM agents share two structural weaknesses. First, they all attempt to prevent attacks

agentsarxiv-cs-cl
13 May 2026
Safety

Emergent Communication between Heterogeneous Visual Agents through Decentralized Learning

DGX agent

arXiv:2605.11695v1 Announce Type: new Abstract: Symbols are shared, but perception is private. We study emergent communication between heterogeneous visual agents through decentralized learning, askin

safetyarxiv-cs-cv
13 May 2026
Model Releases

LongMemEval-V2: Evaluating Long-Term Agent Memory Toward Experienced Colleagues

DGX agent

arXiv:2605.12493v1 Announce Type: new Abstract: Long-term memory is crucial for agents in specialized web environments, where success depends on recalling interface affordances, state dynamics, workfl

model-releasesarxiv-cs-cl
13 May 2026
Agents

MCPShield: Content-Aware Attack Detection for LLM Agent Tool-Call Traffic

DGX agent

arXiv:2605.11053v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has become a widely adopted interface for LLM agents to invoke external tools, yet learned monitoring of MCP tool-cal

agentsarxiv-cs-lg
13 May 2026
Agents

Ace-Skill: Bootstrapping Multimodal Agents with Prioritized and Clustered Evolution

DGX agent

arXiv:2605.08887v1 Announce Type: new Abstract: Self-evolving agents present a promising path toward continual adaptation by distilling task interactions into reusable knowledge artifacts. In practice

agentsarxiv-cs-ai
12 May 2026
Model Releases

AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators

DGX agent

arXiv:2605.08647v1 Announce Type: cross Abstract: Multi-agent systems achieve state-of-the-art outcomes through peer collaboration. However, when an agent in the pipeline silently drops a constraint,

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

AgentForesight: Online Auditing for Early Failure Prediction in Multi-Agent Systems

DGX agent

arXiv:2605.08715v1 Announce Type: cross Abstract: LLM-based multi-agent systems are increasingly deployed on long-horizon tasks, but a single decisive error is often accepted by downstream agents and

model-releasesarxiv-cs-ai
12 May 2026
Agents

Bridging the Cognitive Gap: A Unified Memory Paradigm for 6G Agentic AI-RAN

DGX agent

arXiv:2605.10036v1 Announce Type: cross Abstract: As 6G evolves, the radio access network must transcend traditional automation to embrace agentic AI capable of perception, reasoning, and evolution. A

agentsarxiv-cs-ai
12 May 2026
Model Releases

DSGBench: A Diverse Strategic Game Benchmark for Evaluating LLM-based Agents in Complex Decision-Making Environments

DGX agent

arXiv:2503.06047v2 Announce Type: replace Abstract: Large language model (LLM)-based agents are increasingly applied to complex strategic environments that demand long-horizon reasoning, multi-agent i

model-releasesarxiv-cs-ai
12 May 2026
Safety

Learning Approximate Nash Equilibria in Cooperative Multi-Agent Reinforcement Learning via Mean-Field Subsampling

DGX agent

arXiv:2603.03759v2 Announce Type: replace-cross Abstract: Many large-scale platforms and networked control systems have a centralized decision maker interacting with a massive population of agents und

safetyarxiv-cs-ai
12 May 2026
Safety

OpenClaw-RL: Train Any Agent Simply by Talking

DGX agent

arXiv:2603.10165v2 Announce Type: replace-cross Abstract: Every agent interaction generates a next-state signal, namely the user reply, tool output, terminal or GUI state change that follows each acti

safetyarxiv-cs-ai
12 May 2026
Agents

PECMAN: Perception-enabled Collaborative Multi-Agent Navigation in Unknown Environments

DGX agent

arXiv:2605.09344v1 Announce Type: new Abstract: Most path planners assume fully known, static environments, assumptions that fail when robots navigate in dynamic and partially observable environments.

agentsarxiv-cs-ro
12 May 2026
Safety

Rethinking Ratio-Based Trust Regions for Policy Optimization in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.09212v1 Announce Type: new Abstract: Centralized training with decentralized execution (CTDE) is a standard framework for cooperative multi-agent policy-gradient reinforcement learning, all

safetyarxiv-cs-lg
12 May 2026
Agents

SkillLens: Adaptive Multi-Granularity Skill Reuse for Cost-Efficient LLM Agents

DGX agent

arXiv:2605.08386v1 Announce Type: new Abstract: Skill libraries have become a practical way for LLM agents to reuse procedural experience across tasks. However, existing systems typically treat skills

agentsarxiv-cs-ai
12 May 2026
Agents

SkillMAS: Skill Co-Evolution with LLM-based Multi-Agent System

DGX agent

arXiv:2605.09341v1 Announce Type: cross Abstract: Large language model (LLM) agent systems are increasingly expected to improve after deployment, but existing work often decouples two adaptation targe

agentsarxiv-cs-cl
12 May 2026
Agents

Temperature and Persona Shape LLM Agent Consensus With Minimal Accuracy Gains in Qualitative Coding

DGX agent

arXiv:2507.11198v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) enable new possibilities for qualitative research at scale, including annotation and qualitative coding of educat

agentsarxiv-cs-ai
12 May 2026
Agents

Through the Lens of Character: Resolving Modality-Role Interference in Multimodal Role-Playing Agent

DGX agent

arXiv:2605.09443v1 Announce Type: cross Abstract: The advancement of Multimodal Large Language Models (MLLMs) has expanded Role-Playing Agents (RPAs) into visually grounded environments. However, huma

agentsarxiv-cs-cl
12 May 2026
Agents

Towards a Virtual Neuroscientist: Autonomous Neuroimaging Analysis via Multi-Agent Collaboration

DGX agent

arXiv:2605.09366v1 Announce Type: new Abstract: Transforming neuroimaging data into clinically actionable biomarkers is a knowledge-intensive and labor-intensive process. Standardized workflows such a

agentsarxiv-cs-ai
12 May 2026
Agents

Active Learning for Communication Structure Optimization in LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.05703v2 Announce Type: replace-cross Abstract: Optimizing the communication structure of large language model based multi-agent systems (LLM-MAS) has been shown to improve downstream perfor

agentsarxiv-cs-ai
11 May 2026
Agents

Alternating Target-Path Planning for Scalable Multi-Agent Coordination

DGX agent

arXiv:2605.07744v1 Announce Type: new Abstract: The concurrent target assignment and pathfinding (TAPF) problem extends multi-agent pathfinding (MAPF) by asking planners to allocate distinct targets a

agentsarxiv-cs-ai
11 May 2026
Agents

Conformal Agent Error Attribution

DGX agent

arXiv:2605.06788v1 Announce Type: new Abstract: When multi-agent systems (MAS) fail, identifying where the decisive error occurred is the first step for automated recovery to an earlier state. Error a

agentsarxiv-cs-lg
11 May 2026
← Previous
1…4142434445…233
Next →