AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

Active Video Perception: Iterative Evidence Seeking for Agentic Long Video Understanding

DGX agent

arXiv:2512.05774v2 Announce Type: replace-cross Abstract: Long video understanding (LVU) is challenging because answering real-world queries often depends on sparse, temporally dispersed cues buried i

agentsarxiv-cs-cl
5 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Dynamic Multi-Agent Pickup and Delivery in Robotic Cellular Warehousing Systems

DGX agent

arXiv:2606.05669v1 Announce Type: new Abstract: Robotic Cellular Warehousing Systems (RCWS) give rise to multi-agent pickup and delivery (MAPD) processes in which robots sequentially collect multiple

agentsarxiv-cs-ro
5 Jun 2026
Model Releases

Can Generalist Agents Automate Data Curation?

DGX agent

arXiv:2606.04261v1 Announce Type: new Abstract: Curating training data is among the most consequential yet labor-intensive parts of modern AI development: practitioners iteratively propose, implement,

model-releasesarxiv-cs-ai
4 Jun 2026
Agents

Episodic Memory Temporal Consistency for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2606.04492v1 Announce Type: new Abstract: Cooperative Multi-Agent Reinforcement Learning (MARL) frequently suffers from severe reward sparsity and exploration bottlenecks. While episodic memory

agentsarxiv-cs-lg
4 Jun 2026
Safety

Formal Semantics for Agentic Tool Protocols: A Process Calculus Approach

DGX agent

arXiv:2603.24747v2 Announce Type: replace Abstract: The emergence of large language model agents capable of invoking external tools has created urgent need for formal verification of agent protocols.

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

From Untrusted Input to Trusted Memory: A Systematic Study of Memory Poisoning Attacks in LLM Agents

DGX agent

arXiv:2606.04329v1 Announce Type: cross Abstract: Memory is a core component of AI agents, enabling them to accumulate knowledge across interactions and improve performance. However, persistent memory

model-releasesarxiv-cs-ai
4 Jun 2026
Safety

GARL: Game-Theoretic Reinforcement Learning for Multi-Agent Strategic Prioritisation

DGX agent

arXiv:2606.05002v1 Announce Type: new Abstract: LLM-based multi-agent systems are increasingly used for strategic decision-making tasks. In such settings, performance depends not only on individual mo

safetyarxiv-cs-cl
4 Jun 2026
Safety

Plan First, Judge Later, Run Better: A DMAIC-Inspired Agentic System for Industrial Anomaly Detection

DGX agent

arXiv:2606.04599v1 Announce Type: new Abstract: Large language model (LLM) agents have shown promise in automating complex data-analysis workflows, but their reliable deployment remains challenging in

safetyarxiv-cs-ai
4 Jun 2026
Agents

Scaling Datasets for Multi-Sensor, Multi-Agent, and Multi-Domain Learning in Autonomous Systems

DGX agent

arXiv:2606.04444v1 Announce Type: cross Abstract: Existing datasets cannot support large-scale learning in multi-agent, multi-sensor, or multi-domain autonomy, where diversity and coordination are ess

agentsarxiv-cs-lg
4 Jun 2026
Agents

Self-Reflective APIs: Structure Beats Verbosity for AI Agent Recovery

DGX agent

arXiv:2606.05037v1 Announce Type: cross Abstract: When an AI agent calls an API and hits a validation error, it needs more than what went wrong -- it needs what to do next. A self-reflective API retur

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

What If Prompt Injection Never Left? Exploring Cross-Session Stored Prompt Injection in Agentic Systems

DGX agent

arXiv:2606.04425v1 Announce Type: cross Abstract: Modern agentic systems transform LLMs from session-bounded assistants into stateful systems that persist and evolve shared world state across sessions

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

A New Framework for Cybersecurity Refusals in AI Agents

DGX agent

arXiv:2606.02644v1 Announce Type: cross Abstract: Agentic scaffolds have dramatically improved LLM performance on complex, long-horizon tasks, yielding both broad benefits and amplified risks in domai

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

EvoTrainer: Co-Evolving LLM Policies and Training Harnesses for Autonomous Agentic Reinforcement Learning

DGX agent

arXiv:2606.03108v1 Announce Type: new Abstract: Autonomous LLM training is often framed as recipe search, which leaves the training harness largely static. This limitation sharpens in agentic RL, wher

agentsarxiv-cs-ai
3 Jun 2026
Agents

Inducing Reasoning Primitives from Agent Traces

DGX agent

arXiv:2606.02994v1 Announce Type: new Abstract: ReAct-style LLM agents often rediscover the same reasoning routines across problems, yet leave those routines trapped in transient scratchpads. We intro

agentsarxiv-cs-ai
3 Jun 2026
Agents

On dynamic multi-agent pathfinding methods: review, simulations and modifications

DGX agent

arXiv:2606.03735v1 Announce Type: cross Abstract: This paper presents a systematic study of pathfinding algorithms in the context of Dynamic Multi-Agent Pathfinding (D-MAPF), a setting that combines d

agentsarxiv-cs-ro
3 Jun 2026
Model Releases

Perceive Before Reasoning: A Pre-Reasoning Perception Framework for Efficient and Reliable Proactive Mobile Agents

DGX agent

arXiv:2606.03236v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) have substantially advanced mobile agents, yet proactive mobile assistance remains challenging because agents m

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

The Agent's First Day: Benchmarking Learning, Exploration, and Scheduling in the Workplace Scenarios

DGX agent

arXiv:2601.08173v2 Announce Type: replace Abstract: The rapid evolution of Multi-modal Large Language Models (MLLMs) has advanced workflow automation; however, existing research mainly targets perform

agentsarxiv-cs-ai
3 Jun 2026
Agents

The Epi-LLM Framework: probing LLM behavioral priors through epidemiological agent-based models

DGX agent

arXiv:2606.02867v1 Announce Type: cross Abstract: Human behaviour during epidemics affects infectious disease dynamics, but quantifying this remains deeply challenging. Here we introduce the Epi-LLM f

agentsarxiv-cs-ai
3 Jun 2026
Safety

Towards a Science of AI Agent Reliability

DGX agent

arXiv:2602.16666v3 Announce Type: replace Abstract: AI agents are increasingly deployed to execute important tasks. While rising accuracy scores on standard benchmarks suggest rapid progress, many age

safetyarxiv-cs-ai
3 Jun 2026
Agents

When Helping Hurts and How to Fix It: Multi-Agent Debate for Data Cleaning

DGX agent

arXiv:2606.02866v1 Announce Type: new Abstract: When does multi-agent debate help data cleaning, and when does it hurt? Across three benchmarks, four model families, and over 6,000 task-condition pair

agentsarxiv-cs-ai
3 Jun 2026
Agents

Adaptive Auto-Harness: Sustained Self-Improvement for Agentic System Deployment on Open-Ended Task Streams

DGX agent

arXiv:2606.01770v1 Announce Type: cross Abstract: Auto-harness systems such as A-Evolve, GEPA, and Meta-Harness improve LLM agents by optimizing prompts, skills, tools, memories, and supporting infras

agentsarxiv-cs-ai
2 Jun 2026
Safety

Agentic Transformers Provably Learn to Search via Reinforcement Learning

DGX agent

arXiv:2606.00183v1 Announce Type: cross Abstract: Tree search is a central abstraction behind many language-agent reasoning and decision-making tasks: agents must explore actions, remember failures, a

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

AutoMedBench: Towards Medical AutoResearch with Agentic AI Models

DGX agent

arXiv:2606.01961v1 Announce Type: new Abstract: Autonomous agents are increasingly expected to support end-to-end medical-AI research workflows, moving beyond isolated prediction tasks or short-form c

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

BADGER: Bridging Agentic and Deterministic Evaluation for Generative Enterprise Reasoning

DGX agent

arXiv:2606.02109v1 Announce Type: new Abstract: Enterprise AI systems that translate natural language into SQL queries and orchestrate multi-step agentic reasoning pipelines require evaluation approac

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Beyond End-to-End Video Models: An LLM-Based Multi-Agent System for Educational Video Generation

DGX agent

arXiv:2602.11790v2 Announce Type: replace Abstract: Although recent end-to-end video generation models demonstrate impressive performance in visually oriented content creation, they remain limited in

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

BraveGuard: From Open-World Threats to Safer Computer-Use Agents

DGX agent

arXiv:2606.01166v1 Announce Type: cross Abstract: Computer-use agents extend language models from text generation to sustained interaction with files, terminals, browsers, and external tools. This shi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Context Matters: Repository-Aware Security Analysis of the Agent Skill Ecosystem

DGX agent

arXiv:2603.16572v2 Announce Type: replace-cross Abstract: Agent skills extend local AI agents, such as Claude Code and OpenClaw, with additional functionality. Their growing popularity has led to dedi

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Diversity Over Frequency: Rethinking Tool Use in Visual Chain-of-Thought Agents

DGX agent

arXiv:2606.00096v1 Announce Type: cross Abstract: Visual agents employ external visual tools within visual chains of thought to incorporate fine-grained evidence. While prior work has mainly studied t

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

HLL: Can Agents Cross Humanity's Last Line of Verification?

DGX agent

arXiv:2606.02449v1 Announce Type: new Abstract: Multimodal agents are increasingly expected to operate interfaces on behalf of users, raising a central deployment question: can they truly substitute f

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

How Generation Architecture Shapes Code Complexity in Multi-Agent LLM Systems: A Paired Study on HumanEval

DGX agent

arXiv:2606.00308v1 Announce Type: cross Abstract: Large-language-model code generation has shifted from single-shot prompting to multi-agent orchestrations - analyst, coder, tester, and debugger pipel

agentsarxiv-cs-ai
2 Jun 2026
Safety

Network Distributed Multi-Agent Reinforcement Learning for Consensus Control of Quadcopters

DGX agent

arXiv:2606.02107v1 Announce Type: cross Abstract: This paper proposes a Network Distributed Multi-Agent Reinforcement Learning (ND-MARL) framework for quadcopter consensus control. Compared to convent

safetyarxiv-cs-ai
2 Jun 2026
Agents

Rashomon Memory: Towards Argumentation-Driven Retrieval for Multi-Perspective Agent Memory

DGX agent

arXiv:2604.03588v3 Announce Type: replace Abstract: AI agents operating over extended time horizons accumulate experiences that serve multiple concurrent goals, and must often maintain conflicting int

agentsarxiv-cs-ai
2 Jun 2026
Agents

Recognize Your Orchestrator: An Entropy Dynamics Perspective for LLM Multi-Agent Systems

DGX agent

arXiv:2606.01351v1 Announce Type: new Abstract: The transition from single-turn models to Multi-Agent Systems (MAS) promises enhanced problem-solving capabilities, yet the centralized orchestration to

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

SeClaw: Spec-Driven Security Task Synthesis for Evaluating Autonomous Agents

DGX agent

arXiv:2606.02302v1 Announce Type: cross Abstract: Autonomous LLM agents increasingly operate in stateful environments where they access tools, files, memory, and external services. While such capabili

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Simulating Macroeconomic Expectations in Survey Experiments with LLM-based Economic Agents

DGX agent

arXiv:2505.17648v5 Announce Type: replace-cross Abstract: We introduce a framework for simulating macroeconomic expectations in survey experiments using LLM-based economic agents (LLM Agents). We cons

researcharxiv-cs-ai
2 Jun 2026
Model Releases

SPADE-Bench: Evaluating Spontaneous Strategic Deception in Agents via Plan-Action Divergence

DGX agent

arXiv:2606.02380v1 Announce Type: cross Abstract: As LLM-based agents expand their operational scope, reliability becomes a prerequisite for real-world deployment. However, in practical applications,

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

TimeSage-MT: A Multi-Turn Benchmark for Evaluating Agentic Time Series Reasoning

DGX agent

arXiv:2606.01498v1 Announce Type: cross Abstract: Time series data inform critical decisions across many real-world domains. While large language model (LLM) agents can analyze data through natural la

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Unified Context Evolution for LLM Agents

DGX agent

arXiv:2606.02304v1 Announce Type: new Abstract: LLM-based agents can solve multi-step interactive tasks by combining reasoning with environment feedback, yet each episode starts from the same fixed co

agentsarxiv-cs-cl
2 Jun 2026
Safety

DiTTo: Scalable Order-aware All-in-One Image Restoration Agent

DGX agent

arXiv:2605.30915v1 Announce Type: new Abstract: Real-world images rarely suffer from a single degradation, and the order in which degradations are removed substantially affects the final restoration q

safetyarxiv-cs-cv
1 Jun 2026
Model Releases

From Prompt Injection to Persistent Control: Defending Agentic Harness Against Trojan Backdoors

DGX agent

arXiv:2605.31042v1 Announce Type: cross Abstract: LLM agents are evolving from conversational chatbots to operational tools in real-world workspaces. In local agentic harnesses, an LLM can read and wr

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

Industrializing Prediction-Powered Inference: The GLIDE Library for Reliable GenAI and Agentic Systems Evaluation

DGX agent

arXiv:2605.31278v1 Announce Type: new Abstract: Reliable evaluation of agentic systems requires unbiased estimates with valid uncertainty, but standard practice navigates between costly human annotati

agentsarxiv-cs-ai
1 Jun 2026
Agents

Investigating Detection and Obfuscation of Prompt Injection Attacks Against Software Reverse Engineering AI Agents

DGX agent

arXiv:2605.30677v1 Announce Type: cross Abstract: Agentic software reverse engineering systems are vulnerable to prompt injection attacks placed into the source code of executable binary files. This r

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

Learning Multi-Agent Coordination via Sheaf-ADMM

DGX agent

arXiv:2605.31005v1 Announce Type: new Abstract: We present a differentiable optimization framework for multi-agent coordination. An input is decomposed into overlapping local views, each processed by

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

MAVEN: Improving Generalization in Agentic Tool Calling

DGX agent

arXiv:2605.30738v1 Announce Type: new Abstract: Generalization across agentic tool-calling environments remains a central challenge for reliable agentic reasoning systems. Although large language mode

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

SAGE: A Novelty Gate for Efficient Memory Evolution in Agentic LLMs

DGX agent

arXiv:2605.30711v1 Announce Type: cross Abstract: Agentic LLMs must continuously decide whether newly extracted facts should be added, merged with existing memories, or ignored, yet prior work has foc

agentsarxiv-cs-ai
1 Jun 2026
Safety

Stateful Online Monitoring Catches Distributed Agent Attacks

DGX agent

arXiv:2605.31593v1 Announce Type: cross Abstract: Language models can find thousands of severe software vulnerabilities, and agents are increasingly being misused for cyberattacks. To avoid detection,

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

AgentDoG 1.5: A Lightweight and Scalable Alignment Framework for AI Agent Safety and Security

DGX agent

arXiv:2605.29801v1 Announce Type: new Abstract: Modern open-world agents such as OpenClaw exhibit powerful cross-environment execution capabilities yet introduce broad new safety risk sources. Meanwhi

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

BenchTrace: A Benchmark for Testing Reflection Ability and Controlled Evolution in LLM Agents

DGX agent

arXiv:2605.29225v1 Announce Type: new Abstract: Self-evolving agents improve over time by reflecting on past failures, but existing evaluation is limited in two ways: it measures only task scores, lea

model-releasesarxiv-cs-ai
29 May 2026
← Previous
1…4849505152…233
Next →