AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Local Ai

Rethinking AI Hardware: A Three-Layer Cognitive Architecture for Autonomous Agents

DGX agent

arXiv:2604.13757v1 Announce Type: new Abstract: The next generation of autonomous AI systems will be constrained not only by model capability, but by how intelligence is structured across heterogeneou

local-aiarxiv-cs-ai
17 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

FieldWorkArena: Agentic AI Benchmark for Real Field Work Tasks

DGX agent

arXiv:2505.19662v3 Announce Type: replace-cross Abstract: This paper introduces FieldWorkArena, a benchmark for agentic AI targeting real-world field work. With the recent increase in demand for agent

model-releasesarxiv-cs-cv
16 Apr 2026
Safety

Learning Probabilistic Responsibility Allocations for Multi-Agent Interactions

DGX agent

arXiv:2604.13128v1 Announce Type: cross Abstract: Human behavior in interactive settings is shaped not only by individual objectives but also by shared constraints with others, such as safety. Underst

safetyarxiv-cs-lg
16 Apr 2026
Safety

A longitudinal health agent framework

DGX agent

arXiv:2604.12019v1 Announce Type: new Abstract: Although artificial intelligence (AI) agents are increasingly proposed to support potentially longitudinal health tasks, such as symptom management, beh

safetyarxiv-cs-ai
15 Apr 2026
Agents

AutoSurrogate: An LLM-Driven Multi-Agent Framework for Autonomous Construction of Deep Learning Surrogate Models in Subsurface Flow

DGX agent

arXiv:2604.11945v1 Announce Type: cross Abstract: High-fidelity numerical simulation of subsurface flow is computationally intensive, especially for many-query tasks such as uncertainty quantification

agentsarxiv-cs-ai
15 Apr 2026
Safety

Beyond Static Sandboxing: Learned Capability Governance for Autonomous AI Agents

DGX agent

arXiv:2604.11839v1 Announce Type: cross Abstract: Autonomous AI agents built on open-source runtimes such as OpenClaw expose every available tool to every session by default, regardless of the task. A

safetyarxiv-cs-ai
15 Apr 2026
Safety

EvoSpark: Endogenous Interactive Agent Societies for Unified Long-Horizon Narrative Evolution

DGX agent

arXiv:2604.12776v1 Announce Type: new Abstract: Realizing endogenous narrative evolution in LLM-based multi-agent systems is hindered by the inherent stochasticity of generative emergence. In particul

safetyarxiv-cs-cl
15 Apr 2026
Agents

Malice in Agentland: Down the Rabbit Hole of Backdoors in the AI Supply Chain

DGX agent

arXiv:2510.05159v4 Announce Type: replace-cross Abstract: While finetuning AI agents on interaction data -- such as web browsing or tool use -- improves their capabilities, it also introduces critical

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

Silo-Bench: A Scalable Environment for Evaluating Distributed Coordination in Multi-Agent LLM Systems

DGX agent

arXiv:2603.01045v2 Announce Type: replace-cross Abstract: Large language models are increasingly deployed in multi-agent systems to overcome context limitations by distributing information across agen

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

The Long-Horizon Task Mirage? Diagnosing Where and Why Agentic Systems Break

DGX agent

arXiv:2604.11978v1 Announce Type: new Abstract: Large language model (LLM) agents perform strongly on short- and mid-horizon tasks, but often break down on long-horizon tasks that require extended, in

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Agentic Video Generation: From Text to Executable Event Graphs via Tool-Constrained LLM Planning

DGX agent

arXiv:2604.10383v1 Announce Type: new Abstract: Existing multi-agent video generation systems use LLM agents to orchestrate neural video generators, producing visually impressive but semantically unre

safetyarxiv-cs-cv
14 Apr 2026
Safety

Beyond Message Passing: A Semantic View of Agent Communication Protocols

DGX agent

arXiv:2604.02369v3 Announce Type: replace-cross Abstract: Agent communication protocols are becoming critical infrastructure for large language model (LLM) systems that must use tools, coordinate with

safetyarxiv-cs-ai
14 Apr 2026
Safety

MARLIN: Multi-Agent Reinforcement Learning Guided by Language-Based Inter-Robot Negotiation

DGX agent

arXiv:2410.14383v4 Announce Type: replace Abstract: Multi-agent reinforcement learning is a key method for training multi-robot systems. Through rewarding or punishing robots over a series of episodes

safetyarxiv-cs-ro
14 Apr 2026
Safety

MGA: Memory-Driven GUI Agent for Observation-Centric Interaction

DGX agent

arXiv:2510.24168v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have significantly advanced GUI agents, yet long-horizon automation remains constrained by two critical bot

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

OccuBench: Evaluating AI Agents on Real-World Professional Tasks via Language World Models

DGX agent

arXiv:2604.10866v1 Announce Type: new Abstract: AI agents are expected to perform professional work across hundreds of occupational domains (from emergency department triage to nuclear reactor safety

model-releasesarxiv-cs-cl
14 Apr 2026
Safety

Skill-SD: Skill-Conditioned Self-Distillation for Multi-turn LLM Agents

DGX agent

arXiv:2604.10674v1 Announce Type: cross Abstract: Reinforcement learning (RL) has been widely used to train LLM agents for multi-turn interactive tasks, but its sample efficiency is severely limited b

safetyarxiv-cs-ai
14 Apr 2026
Safety

Stop Fixating on Prompts: Reasoning Hijacking and Constraint Tightening for Red-Teaming LLM Agents

DGX agent

arXiv:2604.05549v2 Announce Type: replace Abstract: With the widespread application of LLM-based agents across various domains, their complexity has introduced new security threats. Existing red-team

safetyarxiv-cs-cl
14 Apr 2026
Agents

Text-Guided 6D Object Pose Rearrangement via Closed-Loop VLM Agents

DGX agent

arXiv:2604.09781v1 Announce Type: new Abstract: Vision-Language Models (VLMs) exhibit strong visual reasoning capabilities, yet they still struggle with 3D understanding. In particular, VLMs often fai

agentsarxiv-cs-cv
14 Apr 2026
Model Releases

Turing Test on Screen: A Benchmark for Mobile GUI Agent Humanization

DGX agent

arXiv:2604.09574v1 Announce Type: new Abstract: The rise of autonomous GUI agents has triggered adversarial countermeasures from digital platforms, yet existing research prioritizes utility and robust

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Constraint-Aware Corrective Memory for Language-Based Drug Discovery Agents

DGX agent

arXiv:2604.09308v1 Announce Type: new Abstract: Large language models are making autonomous drug discovery agents increasingly feasible, but reliable success in this setting is not determined by any s

safetyarxiv-cs-ai
13 Apr 2026
Model Releases

HiL-Bench (Human-in-Loop Benchmark): Do Agents Know When to Ask for Help?

DGX agent

arXiv:2604.09408v1 Announce Type: new Abstract: Frontier coding agents solve complex tasks when given complete context but collapse when specifications are incomplete or ambiguous. The bottleneck is n

model-releasesarxiv-cs-ai
13 Apr 2026
Agents

Interactive ASR: Towards Human-Like Interaction and Semantic Coherence Evaluation for Agentic Speech Recognition

DGX agent

arXiv:2604.09121v1 Announce Type: cross Abstract: Recent years have witnessed remarkable progress in automatic speech recognition (ASR), driven by advances in model architectures and large-scale train

agentsarxiv-cs-ai
13 Apr 2026
Model Releases

SAGE: A Service Agent Graph-guided Evaluation Benchmark

DGX agent

arXiv:2604.09285v1 Announce Type: new Abstract: The development of Large Language Models (LLMs) has catalyzed automation in customer service, yet benchmarking their performance remains challenging. Ex

model-releasesarxiv-cs-ai
13 Apr 2026
Safety

A Physical Agentic Loop for Language-Guided Grasping with Execution-State Monitoring

DGX agent

arXiv:2604.07395v1 Announce Type: cross Abstract: Robotic manipulation systems that follow language instructions often execute grasp primitives in a largely single-shot manner: a model proposes an act

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

ATBench: A Diverse and Realistic Agent Trajectory Benchmark for Safety Evaluation and Diagnosis

DGX agent

arXiv:2604.02022v2 Announce Type: replace Abstract: Evaluating the safety of LLM-based agents is increasingly important because risks in realistic deployments often emerge over multi-step interactions

model-releasesarxiv-cs-ai
10 Apr 2026
Model Releases

ClawBench: Can AI Agents Complete Everyday Online Tasks?

DGX agent

arXiv:2604.08523v1 Announce Type: new Abstract: AI agents may be able to automate your inbox, but can they automate other routine aspects of your life? Everyday online tasks offer a realistic yet unso

model-releasesarxiv-cs-cl
10 Apr 2026
Agents

EMSDialog: Synthetic Multi-person Emergency Medical Service Dialogue Generation from Electronic Patient Care Reports via Multi-LLM Agents

DGX agent

arXiv:2604.07549v1 Announce Type: new Abstract: Conversational diagnosis prediction requires models to track evolving evidence in streaming clinical conversations and decide when to commit to a diagno

agentsarxiv-cs-cl
10 Apr 2026
Safety

Incorporating Social Awareness into Control of Unknown Multi-Agent Systems: A Real-Time Spatiotemporal Tubes Approach

DGX agent

arXiv:2510.25597v2 Announce Type: replace-cross Abstract: This paper presents a decentralized control framework that incorporates social awareness into multi-agent systems with unknown dynamics to ach

safetyarxiv-cs-ro
10 Apr 2026
Agents

Learning to Search: A Decision-Based Agent for Knowledge-Based Visual Question Answering

DGX agent

arXiv:2604.07146v2 Announce Type: replace Abstract: Knowledge-based visual question answering (KB-VQA) requires vision-language models to understand images and use external knowledge, especially for r

agentsarxiv-cs-cv
10 Apr 2026
Safety

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

DGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

safetyarxiv-cs-cl
10 Apr 2026
Agents

OrgForge: A Multi-Agent Simulation Framework for Verifiable Synthetic Corporate Corpora

DGX agent

arXiv:2603.14997v2 Announce Type: replace Abstract: Building and evaluating enterprise AI systems requires synthetic organizational corpora that are internally consistent, temporally structured, and c

agentsarxiv-cs-cl
10 Apr 2026
Agents

PyFi: Toward Pyramid-like Financial Image Understanding for VLMs via Adversarial Agents

DGX agent

arXiv:2512.14735v2 Announce Type: replace-cross Abstract: This paper proposes PyFi, a novel framework for pyramid-like financial image understanding that enables vision language models (VLMs) to reaso

agentsarxiv-cs-ai
10 Apr 2026
Agents

Blast Radius

DGX agent

arXiv:2608.07440v1 Announce Type: new Abstract: Agentic coding faces growing problems of affordability and wasted tokens. We introduce Blast Radius, a predictive memory management layer that estimates

agentsarxiv-cs-ai
10 Aug 2026
Agents

Knowledge-Centric Self-Improvement

DGX agent

arXiv:2607.19592v1 Announce Type: new Abstract: Self-improving AI systems typically treat the agent as the object that improves, by optimizing prompts, workflows, harnesses, or even the agent's own co

agentsarxiv-cs-ai
23 Jul 2026
Agents

When Regulation Has Memory: Hysteresis and Control Burden in Artificial Agency

DGX agent

arXiv:2606.30975v1 Announce Type: new Abstract: Adaptive agents are usually judged by what they do, but an agent can appear stable while the internal effort required to keep it stable is increasing. T

agentsarxiv-cs-ai
1 Jul 2026
Agents

The Chameleon's Limit: Investigating Persona Collapse and Homogenization in Large Language Models

DGX agent

arXiv:2604.24698v1 Announce Type: new Abstract: Applications based on large language models (LLMs), such as multi-agent simulations, require population diversity among agents. We identify a pervasive

agentsarxiv-cs-cl
28 Apr 2026
Agents

QuantClaw: Precision Where It Matters for OpenClaw

DGX agent

arXiv:2604.22577v1 Announce Type: new Abstract: Autonomous agent systems such as OpenClaw introduce significant efficiency challenges due to long-context inputs and multi-turn reasoning. This results

agentsarxiv-cs-ai
27 Apr 2026
Agents

AdaExplore: Failure-Driven Adaptation and Diversity-Preserving Search for Efficient Kernel Generation

DGX agent

arXiv:2604.16625v1 Announce Type: new Abstract: Recent large language model (LLM) agents have shown promise in using execution feedback for test-time adaptation. However, robust self-improvement remai

agentsarxiv-cs-cl
21 Apr 2026
Agents

M^star: Every Task Deserves Its Own Memory Harness

DGX agent

arXiv:2604.11811v1 Announce Type: cross Abstract: Large language model agents rely on specialized memory systems to accumulate and reuse knowledge during extended interactions. Recent architectures ty

agentsarxiv-cs-ai
15 Apr 2026
Agents

Towards grounded autonomous research: an end-to-end LLM mini research loop on published computational physics

DGX agent

arXiv:2604.12198v1 Announce Type: cross Abstract: Recent autonomous LLM agents have demonstrated end-to-end automation of machine-learning research. Real-world physical science is intrinsically harder

agentsarxiv-cs-ai
15 Apr 2026
Agents

Evaluating Cooperation in LLM Social Groups through Elected Leadership

DGX agent

arXiv:2604.11721v1 Announce Type: cross Abstract: Governing common-pool resources requires agents to develop enduring strategies through cooperation and self-governance to avoid collective failure. Wh

agentsarxiv-cs-ai
14 Apr 2026
Agents

Beyond Memory: A Transactional Continuity Kernel for Long-Lived AI Agents

DGX agent

arXiv:2608.11632v1 Announce Type: cross Abstract: Persistent AI agents accumulate versioned state across long horizons, but storage retention alone does not identify authoritative state. Without an ex

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

CTBench: Evaluating Troubleshooting Capabilities of AI Agents in Realistic Telecom Network Operations

DGX agent

arXiv:2608.12002v1 Announce Type: new Abstract: Agents are increasingly considered for automating network operations and maintenance, where engineers must diagnose network faults, optimize configurati

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

DREAMS: Density Functional Theory Based Research Engine for Agentic Materials Simulation

DGX agent

arXiv:2507.14267v2 Announce Type: replace Abstract: Large language model (LLM) agents can execute long-horizon scientific workflows, but their numerical outputs are difficult to trust: agents lose con

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

InfraBench: Evaluating Infrastructure Agents Across Layers, Lifecycle, and Risk

DGX agent

arXiv:2608.11234v1 Announce Type: new Abstract: Managing modern computing infrastructure has become a steadily harder problem due to the ever-increasing complexity. Recent advances in AI agents create

model-releasesarxiv-cs-ai
13 Aug 2026
Agents

Local verification cannot detect non-transportability: a cohomological theory of context preservation in agentic reasoning

DGX agent

arXiv:2608.11252v1 Announce Type: new Abstract: Agentic AI systems routinely transport conclusions across biological, clinical and financial contexts, and the emerging safeguard is local verification:

agentsarxiv-cs-ai
13 Aug 2026
Agents

Social Chain of Thought: A Multi-Agent Architecture Grounded in Medical Differential Diagnosis Methodology

DGX agent

arXiv:2608.11420v1 Announce Type: new Abstract: Medical diagnostic reasoning is a high-impact use case for LLMs that carries significant implications for the health and wellbeing of users. When OpenAI

agentsarxiv-cs-ai
13 Aug 2026
Safety

Actions Speak Louder than Words: Measuring Cross-Lingual Policy Retention in Tool-Using Agents

DGX agent

arXiv:2608.11110v1 Announce Type: new Abstract: When a tool-using agent is given the same task in a different language, does it still take the same steps? Multilingual evaluation rarely asks: it compa

safetyarxiv-cs-cl
12 Aug 2026
← Previous
1…4344454647…233
Next →