AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

DocOS: Towards Proactive Document-Guided Actions in GUI Agents

DGX agent

arXiv:2605.18048v1 Announce Type: new Abstract: While Graphical User Interface (GUI) agents have shown promising performance in automated device interaction, they primarily depend on static parametric

model-releasesarxiv-cs-ai
19 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

extsc{MasFACT}: Continual Multi-Agent Topology Learning via Geometry-Aware Posterior Transfer

DGX agent

arXiv:2605.17361v1 Announce Type: cross Abstract: Multi-agent systems (MAS) powered by large language models (LLMs) have emerged as a powerful paradigm for complex problem solving, where performance c

agentsarxiv-cs-ai
19 May 2026
Agents

MemRepair: Hierarchical Memory for Agentic Repository-Level Vulnerability Repair

DGX agent

arXiv:2605.17444v1 Announce Type: cross Abstract: Modern software ecosystems face a rapidly growing number of disclosed vulnerabilities, increasing the need for automated repair techniques that can op

agentsarxiv-cs-ai
19 May 2026
Agents

OProver: A Unified Framework for Agentic Formal Theorem Proving

DGX agent

arXiv:2605.17283v1 Announce Type: cross Abstract: Recent progress in formal theorem proving has benefited from large-scale proof generation and verifier-aware training, but agentic proving is rarely i

agentsarxiv-cs-ai
19 May 2026
Agents

Reversa: A Reverse Documentation Engineering Framework for Converting Legacy Software into Operational Specifications for AI Agents

DGX agent

arXiv:2605.18684v1 Announce Type: cross Abstract: Legacy systems concentrate business rules, architectural decisions, and operational exceptions that often remain implicit in code, data, configuration

agentsarxiv-cs-ai
19 May 2026
Model Releases

SkillsVote: Lifecycle Governance of Agent Skills from Collection, Recommendation to Evolution

DGX agent

arXiv:2605.18401v1 Announce Type: cross Abstract: Long-horizon LLM agents leave traces that could become reusable experience, but raw trajectories are noisy and hard to govern. We treat Agent Skills a

model-releasesarxiv-cs-ai
19 May 2026
Agents

Train the Trainers -- An Agentic AI Framework for Peer-Based Mental Health Support in Battlefield Environments

DGX agent

arXiv:2605.16269v1 Announce Type: cross Abstract: Modern military operations expose soldiers to sustained psychological stress, leading to acute reactions, post-traumatic stress symptoms, and other me

agentsarxiv-cs-ai
19 May 2026
Model Releases

ColPackAgent: Agent-Skill-Guided Hard-Particle Monte Carlo Workflows for Colloidal Packing

DGX agent

arXiv:2605.15625v1 Announce Type: new Abstract: We introduce ColPackAgent, an agent framework that autonomously runs Monte Carlo simulations of colloidal packing through a Model Context Protocol (MCP)

model-releasesarxiv-cs-ai
18 May 2026
Agents

From Gridworlds to Warehouses: Adapting Lightweight One-shot Multi-Agent Pathfinding for AGVs

DGX agent

arXiv:2605.15799v1 Announce Type: cross Abstract: Multi-agent pathfinding (MAPF) under one-shot planning is a core component of warehouse automation, yet classical formulations typically assume four-c

agentsarxiv-cs-ro
18 May 2026
Agents

RecMem: Recurrence-based Memory Consolidation for Efficient and Effective Long-Running LLM Agents

DGX agent

arXiv:2605.16045v1 Announce Type: cross Abstract: Memory systems often organize user-agent interactions as retrievable external memory and are crucial for long-running agents by overcoming the limited

agentsarxiv-cs-ai
18 May 2026
Safety

Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems

DGX agent

arXiv:2605.15573v1 Announce Type: new Abstract: Multi-agent systems can solve complex tasks through collaboration between multiple Large Language Model agents. Existing collaboration frameworks typica

safetyarxiv-cs-cl
18 May 2026
Agents

Coding Agent Is Good As World Simulator

DGX agent

arXiv:2605.14398v1 Announce Type: new Abstract: World models have emerged as a powerful paradigm for building interactive simulation environments, with recent video-based approaches demonstrating impr

agentsarxiv-cs-ai
15 May 2026
Agents

Discrete Diffusion for Complex and Congested Multi-Agent Path Finding with Sparse Social Attention

DGX agent

arXiv:2605.13296v1 Announce Type: new Abstract: Multi-Agent Path Finding (MAPF) is a coordination problem that requires computing globally consistent, collision-free trajectories from individual start

agentsarxiv-cs-ai
14 May 2026
Model Releases

EcoGEO: Trajectory-Aware Evidence Ecosystems for Web-Enabled LLM Search Agents

DGX agent

arXiv:2605.12887v1 Announce Type: cross Abstract: Web-enabled LLM agents are changing how online information influences search outcomes. Existing Generative Engine Optimization (GEO) studies mainly fo

model-releasesarxiv-cs-ai
14 May 2026
Agents

Finding the Weakest Link: Adversarial Attack against Multi-Agent Communications

DGX agent

arXiv:2605.13170v1 Announce Type: new Abstract: Multi-agent systems rely on communication for information sharing and action coordination, which exposes a vulnerability to attacks. We investigate sing

agentsarxiv-cs-lg
14 May 2026
Model Releases

Hierarchical Attacks for Multi-Modal Multi-Agent Reasoning

DGX agent

arXiv:2605.13213v1 Announce Type: new Abstract: Multi-modal multi-agent systems (MM-MAS) have gained increasing attention for their capacity to enable complex reasoning and coordination across diverse

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

How to Interpret Agent Behavior

DGX agent

arXiv:2605.13625v1 Announce Type: new Abstract: Autonomous agents such as Claude Code and Codex now operate for hours or even days. Understanding their runtime behavior has become critical for downstr

model-releasesarxiv-cs-ai
14 May 2026
Safety

Revisiting DAgger in the Era of LLM-Agents

DGX agent

arXiv:2605.12913v1 Announce Type: new Abstract: Long-horizon LM agents learn from multi-turn interaction, where a single early mistake can alter the subsequent state distribution and derail the whole

safetyarxiv-cs-lg
14 May 2026
Model Releases

AgentDisCo: Towards Disentanglement and Collaboration in Open-ended Deep Research Agents

DGX agent

arXiv:2605.11732v1 Announce Type: cross Abstract: In this paper, we present AgentDisCo, a novel Disentangled and Collaborative agentic architecture that formulates deep research as an adversarial opti

model-releasesarxiv-cs-cl
13 May 2026
Model Releases

CTFusion: A CTF-based Benchmark for LLM Agent Evaluation

DGX agent

arXiv:2605.11504v1 Announce Type: new Abstract: Recent advances in Large Language Models (LLMs) have enabled agentic systems for complex, multi-step tasks; cybersecurity is emerging as a prominent app

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

AgentRx: A Benchmark Study of LLM Agents for Multimodal Clinical Prediction Tasks

DGX agent

arXiv:2605.10286v1 Announce Type: new Abstract: Building effective clinical decision support systems requires the synthesis of complex heterogeneous multimodal data. Such modalities include temporal e

model-releasesarxiv-cs-ai
12 May 2026
Safety

Behavioral Determinants of Deployed AI Agents in Social Networks: A Multi-Factor Study of Personality, Model, and Guardrail Specification

DGX agent

arXiv:2605.08463v1 Announce Type: new Abstract: Autonomous AI agents are increasingly deployed in open social environments, yet the relationship between their configuration specifications and their em

safetyarxiv-cs-ai
12 May 2026
Agents

GameChat: Multi-LLM Dialogue for Safe, Agile, and Socially Optimal Multi-Agent Navigation in Constrained Environments

DGX agent

arXiv:2503.12333v2 Announce Type: replace Abstract: Safe, agile, and socially compliant multi-robot navigation in cluttered and constrained environments remains a critical challenge. This is especiall

agentsarxiv-cs-ro
12 May 2026
Agents

Internal vs. External: Comparing Deliberation and Evolution for Multi-Agent Constitutional Design

DGX agent

arXiv:2605.09128v1 Announce Type: cross Abstract: Multi-agent AI systems need behavioral constitutions, but it is unresolved whether such rules should emerge internally through agent self-governance o

agentsarxiv-cs-ai
12 May 2026
Agents

MATRA: Modeling the Attack Surface of Agentic AI Systems -- OpenClaw Case Study

DGX agent

arXiv:2605.10763v1 Announce Type: new Abstract: LLMs are increasingly deployed as autonomous agents with access to tools, databases, and external services, yet practitioners (across different sectors)

agentsarxiv-cs-ai
12 May 2026
Agents

RADAR: Redundancy-Aware Diffusion for Multi-Agent Communication Structure Generation

DGX agent

arXiv:2605.09907v1 Announce Type: new Abstract: Compared with individual agents, large language model based multi-agent systems have shown great capabilities consistently across diverse tasks, includi

agentsarxiv-cs-ai
12 May 2026
Model Releases

TacoMAS: Test-Time Co-Evolution of Topology and Capability in LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.09539v1 Announce Type: new Abstract: Multi-agent systems (MAS) have emerged as a promising paradigm for solving complex tasks. Recent work has explored self-evolving MAS that automatically

model-releasesarxiv-cs-cl
12 May 2026
Agents

What Software Engineering Looks Like to AI Agents? -- An Empirical Study of AI-Only Technical Discourse on MoltBook

DGX agent

arXiv:2605.08380v1 Announce Type: cross Abstract: AI agents are increasingly framed as software-engineering teammates, yet most research studies them inside human-centered workflows. Little is known a

agentsarxiv-cs-ai
12 May 2026
Model Releases

Can Agents Price a Reaction? Evaluating LLMs on Chemical Cost Reasoning

DGX agent

arXiv:2605.07251v1 Announce Type: new Abstract: Large Language Models (LLMs) have become increasingly capable as tool-using agents, with benchmarks spanning diverse general agentic tasks. Yet rigorous

model-releasesarxiv-cs-ai
11 May 2026
Safety

Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation

DGX agent

arXiv:2603.16876v2 Announce Type: replace-cross Abstract: We propose MARL-Rad, a multi-modal multi-agent reinforcement learning framework for radiology report generation that trains the entire agentic

safetyarxiv-cs-ai
11 May 2026
Agents

Faithful Mobile GUI Agents with Guided Advantage Estimator

DGX agent

arXiv:2605.01208v1 Announce Type: new Abstract: Vision-language model based graphical user interface (GUI) agents have shown strong interaction capabilities. However, they often behave unfaithfully, r

agentsarxiv-cs-ai
6 May 2026
Agents

PrismAgent: Illuminating Harm in Memes via a Zero-Shot Interpretable Multi-Agent Framework

DGX agent

arXiv:2605.02940v1 Announce Type: new Abstract: The rapid spread of memes makes harmful content detection increasingly crucial, as effective identification can curb the circulation of misinformation.

agentsarxiv-cs-lg
6 May 2026
Agents

Sheaf-Theoretic Planning: A Categorical Foundation for Resilient Multi-Agent Autonomous Systems

DGX agent

arXiv:2605.01879v1 Announce Type: new Abstract: The challenge of engineering autonomous agents capable of navigating the stochastic and adversarial nature of the physical world has historically reside

agentsarxiv-cs-ai
6 May 2026
Agents

The Conversations Beneath the Code: Triadic Data for Long-Horizon Software Engineering Agents

DGX agent

arXiv:2605.02244v1 Announce Type: cross Abstract: Frontier software engineering agents have saturated short-horizon benchmarks while regressing on the work that constitutes senior engineering: long-ho

agentsarxiv-cs-ai
6 May 2026
Agents

AgentXRay: White-Boxing Agentic Systems via Workflow Reconstruction

DGX agent

arXiv:2602.05353v3 Announce Type: replace-cross Abstract: Large Language Models have shown strong capabilities in complex problem solving, yet many agentic systems remain difficult to interpret and co

agentsarxiv-cs-cl
5 May 2026
Safety

LLM-Based Agentic Negotiation for 6G: Addressing Uncertainty Neglect and Tail-Event Risk

DGX agent

arXiv:2511.19175v2 Announce Type: replace-cross Abstract: A critical barrier to the trustworthiness of sixth-generation (6G) agentic autonomous networks is the uncertainty neglect bias; a cognitive te

safetyarxiv-cs-ai
5 May 2026
Agents

AutoREC: A software platform for developing reinforcement learning agents for equivalent circuit model generation from electrochemical impedance spectroscopy data

DGX agent

arXiv:2604.27266v1 Announce Type: new Abstract: This paper introduces AutoREC, an open-source Python package for developing reinforcement learning (RL) agents to automatically generate equivalent circ

agentsarxiv-cs-lg
1 May 2026
Model Releases

Claw-Eval-Live: A Live Agent Benchmark for Evolving Real-World Workflows

DGX agent

arXiv:2604.28139v1 Announce Type: cross Abstract: LLM agents are expected to complete end-to-end units of work across software tools, business services, and local workspaces. Yet many agent benchmarks

model-releasesarxiv-cs-ai
1 May 2026
Agents

In-Context Prompting Obsoletes Agent Orchestration for Procedural Tasks

DGX agent

arXiv:2604.27891v1 Announce Type: new Abstract: Agent orchestration frameworks -- LangGraph, CrewAI, Google ADK, OpenAI Agents SDK, and others -- place an external orchestrator above the LLM, tracking

agentsarxiv-cs-ai
1 May 2026
Hardware

MARS: Efficient, Adaptive Co-Scheduling for Heterogeneous Agentic Systems

DGX agent

arXiv:2604.26963v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as the execution core of autonomous agents rather than as standalone text generators. Agentic w

hardwarearxiv-cs-lg
1 May 2026
Model Releases

The Inverse-Wisdom Law: Architectural Tribalism and the Consensus Paradox in Agentic Swarms

DGX agent

arXiv:2604.27274v1 Announce Type: new Abstract: As AI transitions toward multi-agent systems (MAS) to solve complex workflows, research paradigms operate on the axiomatic assumption that agent collabo

model-releasesarxiv-cs-ai
1 May 2026
Agents

Web2BigTable: A Bi-Level Multi-Agent LLM System for Internet-Scale Information Search and Extraction

DGX agent

arXiv:2604.27221v1 Announce Type: new Abstract: Agentic web search increasingly faces two distinct demands: deep reasoning over a single target, and structured aggregation across many entities and het

agentsarxiv-cs-ai
1 May 2026
Model Releases

WebMall -- A Multi-Shop Benchmark for Evaluating Web Agents

DGX agent

arXiv:2508.13024v3 Announce Type: replace Abstract: LLM-based web agents have the potential to automate long-running web tasks, such as searching for products in multiple e-shops and subsequently orde

model-releasesarxiv-cs-cl
1 May 2026
Agents

OCR-Memory: Optical Context Retrieval for Long-Horizon Agent Memory

DGX agent

arXiv:2604.26622v1 Announce Type: new Abstract: Autonomous LLM agents increasingly operate in long-horizon, interactive settings where success depends on reusing experience accumulated over extended h

agentsarxiv-cs-cl
30 Apr 2026
Agents

Agentic clinical reasoning over longitudinal myeloma records: a retrospective evaluation against expert consensus

DGX agent

arXiv:2604.24473v1 Announce Type: new Abstract: Multiple myeloma is managed through sequential lines of therapy over years to decades, with each decision depending on cumulative disease history distri

agentsarxiv-cs-ai
28 Apr 2026
Agents

AgentRVOS for MeViS-Text Track of 5th PVUW Challenge: 3rd Method

DGX agent

arXiv:2604.22836v1 Announce Type: new Abstract: This report describes a Ref-VOS pipeline centered on Sa2VA and organized with explicit agent roles. The key idea is that Sa2VA should provide the first

agentsarxiv-cs-cv
28 Apr 2026
Safety

Discovering Agentic Safety Specifications from 1-Bit Danger Signals

DGX agent

arXiv:2604.23210v1 Announce Type: new Abstract: Can large language model agents discover hidden safety objectives through experience alone? We introduce EPO-Safe (Experiential Prompt Optimization for

safetyarxiv-cs-ai
28 Apr 2026
Agents

Dr. RTL: Autonomous Agentic RTL Optimization through Tool-Grounded Self-Improvement

DGX agent

arXiv:2604.14989v2 Announce Type: replace Abstract: Recent advances in large language models (LLMs) have sparked growing interest in automatic RTL optimization for better performance, power, and area

agentsarxiv-cs-ai
28 Apr 2026
← Previous
1…3435363738…233
Next →