AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Agents

WebXSkill: Skill Learning for Autonomous Web Agents

DGX agent

arXiv:2604.13318v1 Announce Type: cross Abstract: Autonomous web agents powered by large language models (LLMs) have shown promise in completing complex browser tasks, yet they still struggle with lon

agentsarxiv-cs-cl
16 Apr 2026
Safety

Parallax: Why AI Agents That Think Must Never Act

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.12986v1 Announce Type: cross Abstract: Autonomous AI agents are rapidly transitioning from experimental tools to operational infrastructure, with projections that 80% of enterprise applicat

safetyarxiv-cs-ai
15 Apr 2026
Agents

ANCHOR: Branch-Point Data Generation for GUI Agents

DGX agent

arXiv:2602.07153v2 Announce Type: replace Abstract: End-to-end GUI agents for real desktop environments require large amounts of high-quality interaction data, yet collecting human demonstrations is e

agentsarxiv-cs-ai
14 Apr 2026
Agents

Beyond Fluency: Toward Reliable Trajectories in Agentic IR

DGX agent

arXiv:2604.04269v2 Announce Type: replace Abstract: Information Retrieval is shifting from passive document ranking toward autonomous agentic workflows that operate in multi-step Reason-Act-Observe lo

agentsarxiv-cs-ai
14 Apr 2026
Agents

ClawGUI: A Unified Framework for Training, Evaluating, and Deploying GUI Agents

DGX agent

arXiv:2604.11784v1 Announce Type: cross Abstract: GUI agents drive applications through their visual interfaces instead of programmatic APIs, interacting with arbitrary software via taps, swipes, and

agentsarxiv-cs-ai
14 Apr 2026
Agents

Context Kubernetes: Declarative Orchestration of Enterprise Knowledge for Agentic AI Systems

DGX agent

arXiv:2604.11623v1 Announce Type: new Abstract: We introduce Context Kubernetes, an architecture for orchestrating enterprise knowledge in agentic AI systems, with a prototype implementation and eight

agentsarxiv-cs-ai
14 Apr 2026
Agents

FM-Agent: Scaling Formal Methods to Large Systems via LLM-Based Hoare-Style Reasoning

DGX agent

arXiv:2604.11556v1 Announce Type: cross Abstract: LLM-assisted software development has become increasingly prevalent, and can generate large-scale systems, such as compilers. It becomes crucial to st

agentsarxiv-cs-ai
14 Apr 2026
Agents

From Helpful to Trustworthy: LLM Agents for Pair Programming

DGX agent

arXiv:2604.10300v1 Announce Type: cross Abstract: LLM-based coding agents are increasingly used to generate code, tests, and documentation. Still, their outputs can be plausible yet misaligned with de

agentsarxiv-cs-ai
14 Apr 2026
Agents

Prosociality by Coupling, Not Mere Observation: Homeostatic Sharing in an Inspectable Recurrent Artificial Life Agent

DGX agent

arXiv:2604.10760v1 Announce Type: cross Abstract: Artificial agents can be made to 'help' for many reasons, including explicit social reward, hard-coded prosocial bonuses, or direct access to another

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

The Blind Spot of Agent Safety: How Benign User Instructions Expose Critical Vulnerabilities in Computer-Use Agents

DGX agent

arXiv:2604.10577v1 Announce Type: cross Abstract: Computer-use agents (CUAs) can now autonomously complete complex tasks in real digital environments, but when misled, they can also be used to automat

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

VisionClaw: Always-On AI Agents through Smart Glasses

DGX agent

arXiv:2604.03486v2 Announce Type: replace-cross Abstract: We present VisionClaw, an always-on wearable AI agent that integrates live egocentric perception with agentic task execution. Running on Meta

agentsarxiv-cs-ai
10 Apr 2026
Agents

Zero Shot Coordination for Sparse Reward Tasks with Diverse Reward Shapings

DGX agent

arXiv:2604.25076v1 Announce Type: new Abstract: Many Multi-Agent Reinforcement Learning (MARL) agents fail to adapt properly to cooperating with agents trained with the same objectives but different s

agentsarxiv-cs-lg
29 Apr 2026
Agents

GitSkills: A Dataset of Agent Skills on GitHub

DGX agent

arXiv:2608.10906v1 Announce Type: cross Abstract: An agent skill is a folder containing a SKILL.md file with instructions for a language-model agent, optionally accompanied by scripts and reference fi

agentsarxiv-cs-ai
12 Aug 2026
Agents

MESA:Task-Adaptive Multi-Structure Evidence Selection for Long-Horizon Agent Memory

DGX agent

arXiv:2608.10108v1 Announce Type: new Abstract: Long-horizon agents accumulate trajectories spanning hundreds of interleaved reasoning, action, and observation steps, where answering a query may depen

agentsarxiv-cs-ai
12 Aug 2026
Agents

Recovering Wasted Compute in Autoresearch Agents

DGX agent

arXiv:2608.10424v1 Announce Type: new Abstract: A slew of recent works develop agents for solving research problems end-to-end, a paradigm increasingly referred to as autoresearch. Such agents have in

agentsarxiv-cs-ai
12 Aug 2026
Safety

SBCO: Self-Supervised, Verifier-Grounded Harness Optimization For Planning Agents

DGX agent

arXiv:2608.10157v1 Announce Type: new Abstract: Self-improving agents seek to reduce the human engineering effort behind AI systems by enabling them to evolve and self-improve their performance over t

safetyarxiv-cs-ai
12 Aug 2026
Agents

Self-evolving Agentic Customer Support System at LinkedIn

DGX agent

arXiv:2608.10224v1 Announce Type: new Abstract: Enterprise support agents operate in rapidly changing environments where policies, product capabilities, and knowledge bases evolve continuously, making

agentsarxiv-cs-ai
12 Aug 2026
Safety

The CASE Framework: A Multi-Disciplinary Control Architecture for Governing Enterprise Agentic AI

DGX agent

arXiv:2608.10153v1 Announce Type: new Abstract: Enterprises are deploying autonomous AI agents faster than they can govern them, and prevailing approaches stretch a single discipline, typically DevSec

safetyarxiv-cs-ai
12 Aug 2026
Agents

Population-Scalable Multi-Agent World Modeling

DGX agent

arXiv:2608.08600v1 Announce Type: cross Abstract: World models have recently achieved impressive progress in visual prediction and interactive generation, but extending them to multi-agent environment

agentsarxiv-cs-ai
11 Aug 2026
Agents

Thinking Is Not Telling: Information Disclosure in User-Service LLM Agents

DGX agent

arXiv:2602.07796v2 Announce Type: replace Abstract: User-engaged LLM agents increasingly operate in service scenarios where task success depends on coordination between the agent, the user, and a stat

agentsarxiv-cs-cl
11 Aug 2026
Safety

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

DGX agent

arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad

safetyarxiv-cs-ai
11 Aug 2026
Agents

AgentPatch: Coarse-to-Fine Weak-Task Repair for Merging Agentic Multimodal Large Language Models

DGX agent

arXiv:2608.06699v1 Announce Type: new Abstract: Agentic multimodal large language models (MLLMs) extend multimodal perception and reasoning with planning, tool use, and interaction in dynamic environm

agentsarxiv-cs-ai
10 Aug 2026
Local Ai

SAT-Edge-Agent: Hardware-in-the-Loop Edge-Agent Orchestration for Onboard Satellite Intelligence

DGX agent

arXiv:2608.03728v1 Announce Type: new Abstract: Onboard satellite intelligence requires a task layer that translates mission intent into local tool calls, exposes execution state, and returns machine-

local-aiarxiv-cs-ai
5 Aug 2026
Agents

Credit Assignment and Efficient Exploration based on Influence Scope in Multi-agent Reinforcement Learning

DGX agent

arXiv:2505.08630v2 Announce Type: replace Abstract: Training cooperative agents in sparse-reward scenarios poses significant challenges for multi-agent reinforcement learning (MARL). Without clear fee

agentsarxiv-cs-lg
4 Aug 2026
Model Releases

Autonomous Repair for Multi-Agent Systems via Monte-Carlo Tree Search

DGX agent

arXiv:2607.29055v1 Announce Type: cross Abstract: Multi-agent systems (MAS) are increasingly deployed to solve complex tasks. In case of incorrect or unsatisfactory outputs, users have to manually loc

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

MerchantBench: Benchmarking LLM Agents for Long-Term Coherence in E-Commerce Operations

DGX agent

arXiv:2607.28956v1 Announce Type: new Abstract: Large language model agents are increasingly evaluated as autonomous tool users, yet most benchmarks focus on bounded tasks with immediate success crite

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

ChronoMem: Version Control and Semantic Rollback for Large Language Model Agent Memory

DGX agent

arXiv:2607.27773v1 Announce Type: new Abstract: LLM agents increasingly rely on long-term memory to support multi-session interaction and personalization. However, existing agent memory systems are de

model-releasesarxiv-cs-cl
31 Jul 2026
Agents

Do Latent Channels Actually Communicate? A Causal Audit of Latent Multi-Agent LLM

DGX agent

arXiv:2607.26773v1 Announce Type: new Abstract: Latent communication in large language model (LLM)-based multi-agent systems (MAS) transmits continuous internal representations instead of text, but gr

agentsarxiv-cs-ai
31 Jul 2026
Agents

Can AI agents conduct open-ended AI research? Early evidence from two case studies

DGX agent

arXiv:2607.27191v1 Announce Type: cross Abstract: Forecasts of explosive AI progress hinge on AI agents automating AI research. But evidence on whether agents can carry out open-ended AI research is t

agentsarxiv-cs-lg
30 Jul 2026
Agents

(Im)Paired Programming: Coding Agents Improve Productivity but Harm Understanding

DGX agent

arXiv:2607.26375v1 Announce Type: new Abstract: Coding agents (e.g., Cursor) improve developer productivity by optimizing task completion, but shifting users from writing code to prompting and reviewi

agentsarxiv-cs-cl
30 Jul 2026
Model Releases

Two Calls Beat Five Agents: Evaluating Multi-Agent Pipelines Against Self-Refinement for Local Language Models

DGX agent

arXiv:2607.26922v1 Announce Type: new Abstract: Multi-agent LLM pipeline systems break down the task among multiple roles for better reasoning, but are benchmarked mainly with large-scale commercial m

model-releasesarxiv-cs-lg
30 Jul 2026
Agents

How Affect Propagates among LLM Agents: Emergent Emotional Contagion in Crowd Simulation

DGX agent

arXiv:2607.25140v1 Announce Type: new Abstract: This paper studies the behavior of language models in a multi-agent crowd simulation, focusing on how affect propagates among agents that perceive and a

agentsarxiv-cs-ai
29 Jul 2026
Agents

A New Role for Relevance: Guiding Corpus Interaction in Agentic Search

DGX agent

arXiv:2607.24223v1 Announce Type: new Abstract: Relevance is a query-dependent estimate of whether a document or excerpt contains useful evidence. Existing retrieval agents use relevance to select top

agentsarxiv-cs-cl
28 Jul 2026
Agents

Beyond Success Rate: Cost-Aware Evaluation of Offensive and Defensive Security Agents

DGX agent

arXiv:2607.15263v3 Announce Type: replace-cross Abstract: Security-agent evaluations commonly measure peak offensive capability under generous inference budgets, emphasizing vulnerability discovery, e

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

DeepLens Diagnosis Agent: Agentic Workflow Design Lets a Small Reasoning Model Compete with Frontier LLMs

DGX agent

arXiv:2607.22555v1 Announce Type: new Abstract: Medical diagnosis is a multi-stage process: extract facts, consult knowledge, generate a differential analysis, and select the best diagnosis with expla

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

EviBack: Search-Agent Reinforcement Learning via Evidence-Constrained Teacher Backoff

DGX agent

arXiv:2607.23955v1 Announce Type: new Abstract: Reinforcement learning enables Agentic RAG systems to learn multi-turn search from verifiable outcome rewards, but all- zero rollout groups provide no c

agentsarxiv-cs-ai
28 Jul 2026
Safety

Execution-Grounded Security Testing for Coding Agents in Software Engineering Pipelines

DGX agent

arXiv:2607.22569v1 Announce Type: new Abstract: Coding agents are increasingly integrated into system operations, where their tool use can directly modify project artifacts, execution environments, an

safetyarxiv-cs-ai
28 Jul 2026
Agents

How Do Practitioners Build SE Agents? Insights from a Mixed-Methods Study

DGX agent

arXiv:2607.10856v2 Announce Type: replace-cross Abstract: The rise of Software Engineering (SE) agents, i.e., LLM-based agents that can understand large codebases and carry out engineering tasks with

agentsarxiv-cs-ai
28 Jul 2026
Agents

ReasFlow: Assisting Reasoning-Centric Scientific Discovery in Applied Mathematics via a Knowledge-Based Multi-Agent System

DGX agent

arXiv:2607.14178v2 Announce Type: replace Abstract: Recent advances in Large Language Models have fueled autonomous AI agents capable of tackling complex scientific tasks, yet existing automated resea

agentsarxiv-cs-ai
28 Jul 2026
Agents

SF-AMS: Strategic Forgetting for Structured Memory in LLM Agent

DGX agent

arXiv:2607.22562v1 Announce Type: new Abstract: Managing long-context dependencies remains a primary bottleneck in LLM agents, as redundant and irrelevant information can degrade multi-step reasoning.

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Toward User-Conditioned Evaluation of Personal LLM Agents under Temporal Interventions

DGX agent

arXiv:2607.21635v1 Announce Type: new Abstract: Personal agents maintain memories, learned skills, tool configurations, and policy state that evolve with each user. Existing agent benchmarks often eva

model-releasesarxiv-cs-lg
27 Jul 2026
Agents

Streaming Multi-Agent Autoregressive Diffusion Model with World State Registers

DGX agent

arXiv:2607.21594v1 Announce Type: new Abstract: Multi-agent interactive world models should not only generate consistent observations, but also maintain world states that persist across agents and evo

agentsarxiv-cs-cv
24 Jul 2026
Agents

A Survey on Hypergame Theory: Modelling Misaligned Perceptions and Nested Beliefs for Multi-Agent Systems

DGX agent

arXiv:2507.19593v3 Announce Type: replace Abstract: Classical game-theoretic models typically assume rational agents, complete information, and common knowledge of payoffs - assumptions that are often

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

Do Agent Optimizers Compound? A Continual-Learning Evaluation on Terminal-Bench 2.0

DGX agent

arXiv:2607.14004v1 Announce Type: new Abstract: Most reported gains from agent-optimization methods are one-shot: an agent is optimized against a fixed benchmark and the resulting improvement is repor

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

AgentCheck: A Reproduce-Intervene-Mitigate Workbench for LLM Agents over MCP

DGX agent

arXiv:2607.11098v2 Announce Type: replace-cross Abstract: Tool-using LLM agents are mostly evaluated assuming all tools work. When a tool times out, returns a week-stale value, or has its description

agentsarxiv-cs-ai
15 Jul 2026
Agents

Agents That Teach: Towards Designing Incidental Learning Back into AI-Assisted Software Development

DGX agent

arXiv:2607.06101v1 Announce Type: cross Abstract: AI coding agents are rapidly reshaping how software is built, with developers increasingly delegating substantial coding tasks to autonomous agents in

agentsarxiv-cs-ai
8 Jul 2026
Agents

Explicit Credit Assignment through Local Rewards and Dependence Graphs in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2601.21523v2 Announce Type: replace Abstract: To promote cooperation in Multi-Agent Reinforcement Learning, the reward signals of all agents can be aggregated together, forming global rewards th

agentsarxiv-cs-lg
7 Jul 2026
Safety

No Time Like the Present: Agentic Test-Time Training for LLM Agents

DGX agent

arXiv:2607.03441v1 Announce Type: cross Abstract: LLM agents often degrade over long episodes: as trajectories grow, they revisit explored states, repeat failed actions, and lose strategies that previ

safetyarxiv-cs-ai
7 Jul 2026
← Previous
1…910111213…230
Next →