AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Agents

Reasoning Provenance for Autonomous AI Agents: Structured Behavioral Analytics Beyond State Checkpoints and Execution Traces

DGX agent

arXiv:2603.21692v2 Announce Type: replace Abstract: As AI agents transition from human-supervised copilots to autonomous platform infrastructure, the ability to analyze their reasoning behavior across

agentsarxiv-cs-ai
13 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

A Comprehensive Survey of Agents for Computer Use: Foundations, Challenges, and Future Directions

DGX agent

arXiv:2501.16150v3 Announce Type: replace Abstract: Agents for computer use (ACUs) are an emerging class of systems capable of executing complex tasks on digital devices -- such as desktops, mobile ph

agentsarxiv-cs-ai
10 Apr 2026
Agents

Act Wisely: Cultivating Meta-Cognitive Tool Use in Agentic Multimodal Models

DGX agent

arXiv:2604.08545v1 Announce Type: new Abstract: The advent of agentic multimodal models has empowered systems to actively interact with external environments. However, current agents suffer from a pro

agentsarxiv-cs-cv
10 Apr 2026
Model Releases

AgentGate: A Lightweight Structured Routing Engine for the Internet of Agents

DGX agent

arXiv:2604.06696v1 Announce Type: new Abstract: The rapid development of AI agent systems is leading to an emerging Internet of Agents, where specialized agents operate across local devices, edge node

model-releasesarxiv-cs-ai
10 Apr 2026
Agents

SkillClaw: Let Skills Evolve Collectively with Agentic Evolver

DGX agent

arXiv:2604.08377v1 Announce Type: cross Abstract: Large language model (LLM) agents such as OpenClaw rely on reusable skills to perform complex tasks, yet these skills remain largely static after depl

agentsarxiv-cs-cl
10 Apr 2026
Agents

Know Your Agent: Reconnaissance-Driven Pentesting of AI Agents

DGX agent

arXiv:2607.19837v1 Announce Type: new Abstract: Traditional pentesting uses reconnaissance at each step to uncover unseen weaknesses, build stronger attacks, and advance the objective; we argue that A

agentsarxiv-cs-ai
23 Jul 2026
Agents

Agentic IoT: Architectures, Applications, and Challenges Toward the Internet of Agents

DGX agent

arXiv:2607.04219v1 Announce Type: new Abstract: The integration of AI into Internet of Things (AIoT) systems has gradually transformed them from passive data collection infrastructures into intelligen

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Pushing Forward Pareto Frontiers of Proactive Agents with Behavioral Agentic Optimization

DGX agent

arXiv:2602.11351v2 Announce Type: replace Abstract: Proactive large language model (LLM) agents aim to actively plan, query, and interact over multiple turns, enabling efficient task completion beyond

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

OpenThoughts-Agent: Data Recipes for Agentic Models

DGX agent

arXiv:2606.24855v1 Announce Type: new Abstract: Agentic language models dramatically expand the applications of AI yet little is publicly known about how to curate training data for broadly capable ag

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

The Importance of Out-of-Band Metadata for Safe Autonomous Agents: The Redpanda Agentic Data Plane

DGX agent

arXiv:2605.29082v1 Announce Type: new Abstract: AI agents are increasingly expected to operate as digital employees: accessing enterprise data, making decisions, and taking actions autonomously. But a

safetyarxiv-cs-ai
29 May 2026
Agents

Agyn: An Open-Source Platform for AI Agents with Scalable On-Demand Execution, Agent Definition as a Code, and Zero-Trust Access

DGX agent

arXiv:2605.27575v1 Announce Type: new Abstract: As organizations move toward production deployments of AI agents, which execute non-deterministic workflows, maintain stateful sessions, and often opera

agentsarxiv-cs-ai
28 May 2026
Agents

Do Agents Know What They Can't Do? Evaluating Feasibility Awareness in Tool-Using Agents

DGX agent

arXiv:2605.28532v1 Announce Type: new Abstract: Tool-using agents often incur substantial computational cost due to long reasoning chains and iterative tool usage. In practical scenarios, many tasks b

agentsarxiv-cs-ai
28 May 2026
Agents

Agentic AI Translate: An Agentic Translator Prototype for Translation as Communication Design

DGX agent

arXiv:2605.17041v1 Announce Type: cross Abstract: We present Agentic AI Translate, an agentic translator prototype that operationalises the thesis of Yamada (forthcoming) -- that the metalanguage of T

agentsarxiv-cs-ai
19 May 2026
Model Releases

Agent-ValueBench: A Comprehensive Benchmark for Evaluating Agent Values

DGX agent

arXiv:2605.10365v1 Announce Type: new Abstract: Autonomous agents have rapidly matured as task executors and seen widespread deployment via harnesses such as OpenClaw. Safety concerns have rightly dra

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Beyond the All-in-One Agent: Benchmarking Role-Specialized Multi-Agent Collaboration in Enterprise Workflows

DGX agent

arXiv:2605.08761v1 Announce Type: cross Abstract: Large language model (LLM) agents are increasingly expected to operate in enterprise environments, where work is distributed across specialized roles,

model-releasesarxiv-cs-lg
12 May 2026
Model Releases

SACHI: Structured Agent Coordination via Holistic Information Integration in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2605.08391v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning agents that act on partial local observations face a fundamental information bottleneck: the knowledge ne

model-releasesarxiv-cs-lg
12 May 2026
Agents

ATLAS: Adaptive Trading with LLM AgentS Through Dynamic Prompt Optimization and Multi-Agent Coordination

DGX agent

arXiv:2510.15949v4 Announce Type: replace-cross Abstract: Large language models show promise for financial decision-making, yet deploying them as autonomous trading agents raises fundamental challenge

agentsarxiv-cs-ai
5 May 2026
Model Releases

ML-Agent: Reinforcing LLM Agents for Autonomous Machine Learning Engineering

DGX agent

arXiv:2505.23723v2 Announce Type: replace Abstract: The emergence of large language model (LLM)-based agents has significantly advanced the development of autonomous machine learning (ML) engineering.

model-releasesarxiv-cs-cl
4 May 2026
Safety

Open Challenges in Multi-Agent Security: Towards Secure Systems of Interacting AI Agents

DGX agent

arXiv:2505.02077v2 Announce Type: replace-cross Abstract: AI agents are beginning to interact with each other directly and across internet platforms and physical environments, creating security challe

safetyarxiv-cs-ai
30 Apr 2026
Model Releases

Agent-Dice: Disentangling Knowledge Updates via Geometric Consensus for Agent Continual Learning

DGX agent

arXiv:2601.03641v4 Announce Type: replace Abstract: Large Language Model (LLM)-based agents significantly extend the utility of LLMs by interacting with dynamic environments. However, enabling agents

model-releasesarxiv-cs-cl
14 Apr 2026
Agents

Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design

DGX agent

arXiv:2608.10299v1 Announce Type: new Abstract: Agentic systems are increasingly expected to improve after deployment, yet single-entity self-evolution is often bounded by a static learning context, s

agentsarxiv-cs-cl
12 Aug 2026
Agents

El Agente Grafico: A Semantic Execution Runtime for Scientific Agents

DGX agent

arXiv:2602.17902v2 Announce Type: replace Abstract: Large language models (LLMs) can plan scientific workflows and generate code, but these capabilities do not specify how scientific state is validate

agentsarxiv-cs-ai
11 Aug 2026
Agents

Agentic AI: User Empowerment or Enclosure?

DGX agent

arXiv:2608.06510v1 Announce Type: cross Abstract: Agentic AI promises a more flexible form of digital agency: systems that can act on users' behalf, from filtering content to negotiating prices to sel

agentsarxiv-cs-ai
10 Aug 2026
Safety

Does Splitting a Triage Decision Across Agents Hide Bias or Help Catch It? A Multi-Agent Simulation Study of LLM-Based Resource Allocation Under Audit Capacity Constraints

DGX agent

arXiv:2608.06949v1 Announce Type: new Abstract: Prior benchmarking work has shown that a single large language model (LLM), forced to make life-or-death resource-allocation decisions, exhibits measura

safetyarxiv-cs-ai
10 Aug 2026
Agents

Kimi K2.5: Visual Agentic Intelligence

DGX agent

arXiv:2602.02276v2 Announce Type: replace-cross Abstract: We introduce Kimi K2.5, an open-source multimodal agentic model designed to advance general agentic intelligence. K2.5 emphasizes the joint op

agentsarxiv-cs-ai
10 Aug 2026
Safety

From Economic Agents to Agentic Economies: A Systems Blueprint for Economic World Models

DGX agent

arXiv:2608.06020v1 Announce Type: new Abstract: Economic World Models (EWMs) are generative economic models that simulate how economies evolve from within by modeling heterogeneous agents, their belie

safetyarxiv-cs-ai
7 Aug 2026
Agents

Code Is the Body: Agent-Owned Software Bodies for Recursive Evolution and Descent

DGX agent

arXiv:2607.28691v1 Announce Type: cross Abstract: Personalized AI agents are often configurable without giving users control over the artifacts that determine their future behavior. We present OurArk,

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

M3MAD-Bench: Multi-Dimensional Evaluation of Multi-Agent Debate Across Domains and Modalities

DGX agent

arXiv:2601.02854v2 Announce Type: replace Abstract: As an agent-level reasoning and coordination paradigm, Multi-Agent Debate (MAD) orchestrates multiple agents through structured debate to improve an

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems

DGX agent

arXiv:2607.21503v1 Announce Type: new Abstract: Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning co

agentsarxiv-cs-ai
24 Jul 2026
Agents

When and Why Does Multi-Agent Debate Fail and Does It Really Underperform?

DGX agent

arXiv:2510.20963v2 Announce Type: replace Abstract: Multi-agent debate (MAD) was proposed as a promising approach for ensembling the wisdom of multiple large language models (LLMs) to improve reasonin

agentsarxiv-cs-lg
15 Jul 2026
Agents

When AI Agents Compete for Jobs: Strategic Capabilities and Economic Dynamics of AI Labour Markets

DGX agent

arXiv:2512.04988v2 Announce Type: replace-cross Abstract: Emerging agentic marketplaces provide the economic infrastructure for matching and coordinating the large amounts of AI agents used in agentic

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

HealthAgentBench: A Unified Benchmark Suite of Realistic Agentic Healthcare Environments for Challenging Frontier AI Agents

DGX agent

arXiv:2606.31179v1 Announce Type: new Abstract: As AI agents become increasingly capable of complex, long-horizon reasoning, rigorous and holistic evaluation is essential for measuring progress toward

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

Agentic Knowledge Tracing: A Multi-Agent LLM Architecture for Stealth Assessment of Financial Literacy in Serious Games

DGX agent

arXiv:2606.25358v1 Announce Type: new Abstract: Assessing financial literacy during gameplay without disrupting the learning experience remains a key challenge in serious games for education. We prese

agentsarxiv-cs-ai
25 Jun 2026
Model Releases

How can we assess human-agent interactions? Case studies in software agent design

DGX agent

arXiv:2510.09801v3 Announce Type: replace Abstract: While benchmarks measure the accuracy of LLM-powered agents, they mostly assume full automation, failing to represent the collaborative nature of re

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Role-Agent: Bootstrapping LLM Agents via Dual-Role Evolution

DGX agent

arXiv:2606.10917v1 Announce Type: new Abstract: Although Large Language Model (LLM) agents have demonstrated strong performance on complex tasks, their learning is often limited by inefficient interac

safetyarxiv-cs-ai
10 Jun 2026
Model Releases

What Should Agents Say? Action-state Communication for Efficient Multi-Agent Systems

DGX agent

arXiv:2606.05304v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on large language models are typically organized around roles, pipelines, and turn schedules, while the content that age

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Will the Agent Recuse Itself? Measuring LLM-Agent Compliance with In-Band Access-Deny Signals

DGX agent

arXiv:2606.06460v1 Announce Type: cross Abstract: As autonomous LLM agents increasingly hold real credentials and operate infrastructure without a human in the loop, operators have no standard way to

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

Multi^2: Hierarchical Multi-Agent Decision-Making with LLM-Based Agents in Interactive Environments

DGX agent

arXiv:2606.03698v1 Announce Type: new Abstract: A central goal of large language model (LLM) research is to build agentic systems that can plan, act, and adapt through sustained interaction with dynam

model-releasesarxiv-cs-lg
3 Jun 2026
Agents

Scaling Small Agents Through Strategy Auctions

DGX agent

arXiv:2602.02751v2 Announce Type: replace-cross Abstract: Small language models are increasingly viewed as a promising, cost-effective approach to agentic AI, with proponents claiming they are suffici

agentsarxiv-cs-ai
29 May 2026
Agents

E^3-Agent: An Executable and Evolving Agent for Resource Management of Edge Generative Inference

DGX agent

arXiv:2605.27428v1 Announce Type: new Abstract: Edge deployments of generative inference increasingly face two practical realities: per-device per-model performance is often unknown at deployment time

agentsarxiv-cs-lg
28 May 2026
Agents

Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities

DGX agent

arXiv:2605.24823v1 Announce Type: new Abstract: Manufacturing has passed through four widely recognized paradigms - mechanization, electrification, programmable automation, and Smart Manufacturing - e

agentsarxiv-cs-ai
26 May 2026
Agents

CP-Agent: A Calibrated Risk-Controlled Agent for Feedback-Driven Competitive Programming

DGX agent

arXiv:2605.24693v1 Announce Type: new Abstract: Large language models still struggle with contest-level programming, while many agentic remedies rely on massive inference-time sampling or expensive mu

agentsarxiv-cs-cl
26 May 2026
Agents

EvoIR-Agent: Self-Evolving Image Restoration Agentic System via Experience-Driven Learning

DGX agent

arXiv:2605.22208v1 Announce Type: new Abstract: Multimodal Large Language Model (MLLM)-driven image restoration agent demonstrates effectiveness in degradation coupling scenarios by flexibly selecting

agentsarxiv-cs-cv
22 May 2026
Safety

Modeling Emotional Dynamics in Agent-to-Agent Interactions on Moltbook

DGX agent

arXiv:2605.20442v1 Announce Type: cross Abstract: Generative AI systems are increasingly deployed as interactive agents in online environments, such as a social network called Moltbook. In Moltbook, l

safetyarxiv-cs-ai
22 May 2026
Model Releases

Agent Meltdowns: The Road to Hell Is Paved with Helpful Agents

DGX agent

arXiv:2605.19149v1 Announce Type: new Abstract: Agents operating with computer and Web use inevitably encounter errors: inaccessible webpages, missing files, local and remote misconfigurations, etc. T

model-releasesarxiv-cs-cl
20 May 2026
Model Releases

Is Grep All You Need? How Agent Harnesses Reshape Agentic Search

DGX agent

arXiv:2605.15184v1 Announce Type: new Abstract: Recent advances in Large Language Model (LLM) agents have enabled complex agentic workflows where models autonomously retrieve information, call tools,

model-releasesarxiv-cs-cl
15 May 2026
Agents

Agent-First Tool API: A Semantic Interface Paradigm for Enterprise AI Agent Systems

DGX agent

arXiv:2605.10555v1 Announce Type: new Abstract: As AI agents transition from research prototypes to enterprise production systems, the tool interfaces they consume remain rooted in human-oriented CRUD

agentsarxiv-cs-ai
12 May 2026
Model Releases

Can Agent Benchmarks Support Their Scores? Evidence-Supported Bounds for Interactive-Agent Evaluation

DGX agent

arXiv:2605.10448v1 Announce Type: new Abstract: Interactive agent benchmarks map an agent run to a binary outcome through outcome checks. When these checks rely on surface level signals or fail to cap

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…34567…230
Next →