AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Agents

Provably Auditable and Safe LLM Agents from Human-Authored Ontologies

DGX agent

arXiv:2606.04903v1 Announce Type: cross Abstract: We introduce the LLM agent architecture Agentic Redux, intended for use with nontrivial problem domains that require linear auditability. Using the ty

agentsarxiv-cs-ai
4 Jun 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SePO: Self-Evolving Prompt Agent for System Prompt Optimization

DGX agent

arXiv:2606.04465v1 Announce Type: cross Abstract: System prompt optimization improves agent behavior without modifying the underlying model, yielding human-readable, model-agnostic instructions. Exist

agentsarxiv-cs-ai
4 Jun 2026
Agents

Generative Multi-Robot Motion Planning via Diffusion Modeling with Multi-Agent Reinforcement Learning Guidance

DGX agent

arXiv:2606.00933v1 Announce Type: new Abstract: Coordinating multiple robots in shared environments requires generating feasible trajectories for each agent while accounting for interactions among age

agentsarxiv-cs-ro
2 Jun 2026
Agents

Learning to Construct Practical Agentic Systems

DGX agent

arXiv:2606.00189v1 Announce Type: cross Abstract: Automated design and optimization of agentic LLM-based systems leads to sophisticated systems that substantially improve result quality over off-the-s

agentsarxiv-cs-ai
2 Jun 2026
Agents

Counterfactual Graph for Multi-Agent LLM Calibration

DGX agent

arXiv:2605.30653v1 Announce Type: new Abstract: Multi-agent LLM systems often treat agreement as evidence: when many agents in a panel give the same answer, that answer is assumed to be more reliable.

agentsarxiv-cs-cl
1 Jun 2026
Model Releases

Battery-Sim-Agent: Leveraging LLM-Agent for Inverse Battery Parameter Estimation

DGX agent

arXiv:2605.29560v1 Announce Type: new Abstract: Parameterizing high-fidelity 'digital twins' of batteries is a critical yet challenging inverse problem that hinders the pace of battery innovation. Pre

model-releasesarxiv-cs-ai
29 May 2026
Agents

CONCAT: Consensus- and Confidence-Driven Ad Hoc Teaming for Efficient LLM-Based Multi-Agent Systems

DGX agent

arXiv:2605.29612v1 Announce Type: cross Abstract: Although large language model (LLM) based multi-agent systems (MAS) show their capability to solve complex tasks and achieve higher performance over s

agentsarxiv-cs-cl
29 May 2026
Safety

SafeRx-Agent: A Knowledge-Grounded Multi-Agent Framework for Safe and Explainable Medication Recommendation

DGX agent

arXiv:2605.29146v1 Announce Type: cross Abstract: Medication recommendation predicts medications for patient visits, but existing methods still face two key challenges. At the model level, traditional

safetyarxiv-cs-ai
29 May 2026
Safety

Colosseum: Auditing Collusion in Cooperative Multi-Agent Systems

DGX agent

arXiv:2602.15198v2 Announce Type: replace-cross Abstract: Multi-agent systems, where LLM agents communicate through free-form language, enable sophisticated coordination for solving complex cooperativ

safetyarxiv-cs-ai
28 May 2026
Local Ai

Is Agent Memory a Database? Rethinking Data Foundations for Long-Term AI Agent Memory

DGX agent

arXiv:2605.26252v1 Announce Type: new Abstract: Long-running AI agents need persistent memory. Memory supports learning across sessions, reduces repeated context injection, and enables auditing of pas

local-aiarxiv-cs-ai
27 May 2026
Safety

A Sober Look at Agentic Misalignment in Automated Workflows

DGX agent

arXiv:2605.24197v1 Announce Type: new Abstract: We study a class of emergent misalignment in multi-agent systems (MAS), with a focus on automated workflows, which we refer to agentic misalignment. Alt

safetyarxiv-cs-ai
26 May 2026
Agents

6G Communication Networks Enabling Embodied Agents: Architecture and Prototype

DGX agent

arXiv:2605.23263v1 Announce Type: cross Abstract: Embodied agents, which couple intelligent decision-making with physical actuation in the real world, impose far more stringent and heterogeneous commu

agentsarxiv-cs-ai
25 May 2026
Agents

GradingAttack: Exposing Security Vulnerabilities in LLM Based Educational Grading Agents

DGX agent

arXiv:2602.00979v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly deployed as educational agents for automatic short answer grading (ASAG) in real-world education

agentsarxiv-cs-ai
25 May 2026
Safety

Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness

DGX agent

arXiv:2605.23146v1 Announce Type: cross Abstract: Classical reinforcement learning assumes the agent interacts with a fixed environment whose behavior does not depend on the agent's policy. This assum

safetyarxiv-cs-ai
25 May 2026
Model Releases

Terminal-World: Scaling Terminal-Agent Environments via Agent Skills

DGX agent

arXiv:2605.20876v1 Announce Type: new Abstract: Terminal agents extend Large Language Models with the ability to execute tasks directly in command-line environments, but their progress is bottlenecked

model-releasesarxiv-cs-cl
21 May 2026
Model Releases

VerifyMAS: Hypothesis Verification for Failure Attribution in LLM Multi-Agent Systems

DGX agent

arXiv:2605.17467v1 Announce Type: new Abstract: Large language model-driven multi-agent systems (LLM-MAS) excel at complex tasks, yet unreliable agents remain a key bottleneck to system-level reliabil

model-releasesarxiv-cs-cl
19 May 2026
Model Releases

CAX-Agent: A Lightweight Agent Harness for Reliable APDL Automation

DGX agent

arXiv:2605.15218v1 Announce Type: new Abstract: Large language models deployed for MAPDL finite-element simulation face practical reliability challenges: without structured execution control, tool enc

model-releasesarxiv-cs-ai
18 May 2026
Agents

Language-Based Agent Control

DGX agent

arXiv:2605.12863v1 Announce Type: cross Abstract: This paper introduces language-based agent control (LBAC), a new programming model for agentic applications that brings techniques from programming la

agentsarxiv-cs-ai
14 May 2026
Model Releases

AHD Agent: Agentic Reinforcement Learning for Automatic Heuristic Design

DGX agent

arXiv:2605.08756v1 Announce Type: new Abstract: Automatic heuristic design (AHD) has emerged as a promising paradigm for solving NP-hard combinatorial optimization problems (COPs). Recent works show t

model-releasesarxiv-cs-ai
12 May 2026
Agents

Evolutionary Ensemble of Agents

DGX agent

arXiv:2605.09018v1 Announce Type: cross Abstract: We introduce Evolutionary Ensemble (EvE), a decentralized framework that organizes existing, highly capable coding agents into a live, co-evolving sys

agentsarxiv-cs-ai
12 May 2026
Model Releases

General Agent Evaluation

DGX agent

arXiv:2602.22953v2 Announce Type: replace Abstract: General-purpose agents perform tasks in unfamiliar environments without domain-specific manual customization. Yet no study has systematically measur

model-releasesarxiv-cs-ai
12 May 2026
Agents

Slipstream: Trajectory-Grounded Compaction Validation for Long-Horizon Agents

DGX agent

arXiv:2605.08580v1 Announce Type: cross Abstract: To cope with the large contexts that long-horizon LLM agents produce, modern frameworks increasingly rely on compaction -- invoking an LLM to rewrite

agentsarxiv-cs-ai
12 May 2026
Agents

Token Economics for LLM Agents: A Dual-View Study from Computing and Economics

DGX agent

arXiv:2605.09104v1 Announce Type: new Abstract: As LLM agents evolve, tokens have emerged as the core economic primitives of Agentic AI. However, their exponential consumption introduces severe comput

agentsarxiv-cs-ai
12 May 2026
Model Releases

Agentick: A Unified Benchmark for General Sequential Decision-Making Agents

DGX agent

arXiv:2605.06869v1 Announce Type: new Abstract: AI agent research spans a wide spectrum: from RL agents that learn from scratch to foundation model agents that leverage pre-trained knowledge, yet no u

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

BioProVLA-Agent: An Affordable, Protocol-Driven, Vision-Enhanced VLA-Enabled Embodied Multi-Agent System with Closed-Loop-Capable Reasoning for Biological Laboratory Manipulation

DGX agent

arXiv:2605.07306v1 Announce Type: cross Abstract: Biological laboratory automation can reduce repetitive manual work and improve reproducibility, but reliable embodied execution in wet-lab environment

model-releasesarxiv-cs-ai
11 May 2026
Agents

Position: Agent Should Invoke External Tools ONLY When Epistemically Necessary

DGX agent

arXiv:2506.00886v3 Announce Type: replace Abstract: As large language models evolve into tool-augmented agents, a central question remains unresolved: when is external tool use actually justified? Exi

agentsarxiv-cs-ai
7 May 2026
Agents

ProMediate: A Socio-cognitive framework for evaluating proactive agents in multi-party negotiation

DGX agent

arXiv:2510.25224v3 Announce Type: replace Abstract: While Large Language Models (LLMs) are increasingly used in agentic frameworks to assist individual users, there is a growing need for agents that c

agentsarxiv-cs-cl
7 May 2026
Model Releases

12 Angry AI Agents: Evaluating Multi-Agent LLM Decision-Making Through Cinematic Jury Deliberation

DGX agent

arXiv:2605.01986v1 Announce Type: new Abstract: What if the twelve jurors of Sidney Lumet's 12 Angry Men (1957) were not men, but large language models? Would the one juror who disagrees still be able

model-releasesarxiv-cs-ai
6 May 2026
Safety

Descent-Guided Policy Gradient for Scalable Cooperative Multi-Agent Learning

DGX agent

arXiv:2602.20078v3 Announce Type: replace-cross Abstract: Scaling cooperative multi-agent reinforcement learning (MARL) is fundamentally limited by cross-agent noise. When agents share a common reward

safetyarxiv-cs-lg
6 May 2026
Agents

LLM-Powered AI Agent Systems and Their Applications in Industry

DGX agent

arXiv:2505.16120v2 Announce Type: replace Abstract: The emergence of Large Language Models (LLMs) has reshaped agent systems. Unlike traditional rule-based agents with limited task scope, LLM-powered

agentsarxiv-cs-ai
6 May 2026
Agents

Quality-Aware Exploration Budget Allocation for Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.01865v1 Announce Type: cross Abstract: Cooperative multi-agent reinforcement learning (MARL) requires agents to discover joint strategies in a combinatorially large state-action space, yet

agentsarxiv-cs-ai
6 May 2026
Model Releases

E-mem: Multi-agent based Episodic Context Reconstruction for LLM Agent Memory

DGX agent

arXiv:2601.21714v2 Announce Type: replace Abstract: The evolution of Large Language Model (LLM) agents towards System~2 reasoning, characterized by deliberative, high-precision problem-solving, requir

model-releasesarxiv-cs-ai
5 May 2026
Agents

GRAIL: A Deep-Granularity Hybrid Resonance Framework for Real-Time Agent Discovery via SLM-Enhanced Indexing

DGX agent

arXiv:2605.02489v1 Announce Type: cross Abstract: As the ecosystem of Large Language Model (LLM)-based agents expands rapidly, efficient and accurate Agent Discovery becomes a critical bottleneck for

agentsarxiv-cs-cl
5 May 2026
Agents

Learning to Rewrite Tool Descriptions for Reliable LLM-Agent Tool Use

DGX agent

arXiv:2602.20426v2 Announce Type: replace Abstract: While most efforts to improve LLM-based tool-using agents focus on the agent itself - through larger models, better prompting, or fine-tuning - agen

agentsarxiv-cs-ai
30 Apr 2026
Agents

From Soliloquy to Agora: Memory-Enhanced LLM Agents with Decentralized Debate for Optimization Modeling

DGX agent

arXiv:2604.25847v1 Announce Type: cross Abstract: Optimization modeling underpins real-world decision-making in logistics, manufacturing, energy, and public services, but reliably solving such problem

agentsarxiv-cs-lg
29 Apr 2026
Agents

AgentEval: DAG-Structured Step-Level Evaluation for Agentic Workflows with Error Propagation Tracking

DGX agent

arXiv:2604.23581v1 Announce Type: cross Abstract: Agentic systems that chain reasoning, tool use, and synthesis into multi-step workflows are entering production, yet prevailing evaluation practices l

agentsarxiv-cs-cl
28 Apr 2026
Safety

Security Considerations for Multi-agent Systems

DGX agent

arXiv:2603.09002v2 Announce Type: replace-cross Abstract: Multi-agent artificial intelligence systems or MAS are systems of autonomous agents that exercise delegated tool authority, share persistent m

safetyarxiv-cs-ai
28 Apr 2026
Safety

AdaFair-MARL: Enforcing Adaptive Fairness Constraints in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2511.14135v2 Announce Type: replace-cross Abstract: Fair workload enforcement in heterogeneous multi-agent systems that pursue shared objectives remains challenging. Fixed fairness penalties oft

safetyarxiv-cs-ai
27 Apr 2026
Agents

Efficient Multi-Agent System Training with Data Influence-Oriented Tree Search

DGX agent

arXiv:2502.00955v2 Announce Type: replace Abstract: Monte Carlo Tree Search (MCTS) based methods provide promising approaches for generating synthetic data to enhance the self-training of Large Langua

agentsarxiv-cs-cl
27 Apr 2026
Agents

From Actions to Understanding: Conformal Interpretability of Temporal Concepts in LLM Agents

DGX agent

arXiv:2604.19775v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous agents capable of reasoning, planning, and acting within interactive environments.

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

SkillLearnBench: Benchmarking Continual Learning Methods for Agent Skill Generation on Real-World Tasks

DGX agent

arXiv:2604.20087v1 Announce Type: new Abstract: Skills have become the de facto way to enable LLM agents to perform complex real-world tasks with customized instructions, workflows, and tools, but how

model-releasesarxiv-cs-cl
23 Apr 2026
Agents

LPO: Towards Accurate GUI Agent Interaction via Location Preference Optimization

DGX agent

arXiv:2506.09373v3 Announce Type: replace-cross Abstract: The advent of autonomous agents is transforming interactions with Graphical User Interfaces (GUIs) by employing natural language as a powerful

agentsarxiv-cs-ai
22 Apr 2026
Model Releases

MiroThinker: Pushing the Performance Boundaries of Open-Source Research Agents via Model, Context, and Interactive Scaling

DGX agent

arXiv:2511.11793v3 Announce Type: replace Abstract: We present MiroThinker v1.0, an open-source research agent designed to advance tool-augmented reasoning and information-seeking capabilities. Unlike

model-releasesarxiv-cs-cl
22 Apr 2026
Research

Temporal UI State Inconsistency in Desktop GUI Agents: Formalizing and Defending Against TOCTOU Attacks on Computer-Use Agents

DGX agent

arXiv:2604.18860v1 Announce Type: cross Abstract: GUI agents that control desktop computers via screenshot-and-click loops introduce a new class of vulnerability: the observation-to-action gap (mean 6

researcharxiv-cs-ai
22 Apr 2026
Agents

A Discordance-Aware Multimodal Framework with Multi-Agent Clinical Reasoning

DGX agent

arXiv:2604.16333v1 Announce Type: new Abstract: Knee osteoarthritis frequently exhibits discordance between structural damage observed in imaging and patient-reported symptoms such as pain. This misma

agentsarxiv-cs-lg
21 Apr 2026
Agents

Diversity Collapse in Multi-Agent LLM Systems: Structural Coupling and Collective Failure in Open-Ended Idea Generation

DGX agent

arXiv:2604.18005v1 Announce Type: cross Abstract: Multi-agent systems (MAS) are increasingly used for open-ended idea generation, driven by the expectation that collective interaction will broaden the

agentsarxiv-cs-cl
21 Apr 2026
Agents

DVAR: Adversarial Multi-Agent Debate for Video Authenticity Detection

DGX agent

arXiv:2604.16987v1 Announce Type: new Abstract: The rapid evolution of video generation technologies poses a significant challenge to media forensics, as conventional detection methods often fail to g

agentsarxiv-cs-cv
21 Apr 2026
Model Releases

EmbodiedLGR: Integrating Lightweight Graph Representation and Retrieval for Semantic-Spatial Memory in Robotic Agents

DGX agent

arXiv:2604.18271v1 Announce Type: new Abstract: As the world of agentic artificial intelligence applied to robotics evolves, the need for agents capable of building and retrieving memories and observa

model-releasesarxiv-cs-ro
21 Apr 2026
← Previous
1…1213141516…230
Next →