AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

Semia: Auditing Agent Skills via Constraint-Guided Representation Synthesis

DGX agent

arXiv:2605.00314v1 Announce Type: cross Abstract: An agent skill is a configuration package that equips an LLM-driven agent with a concrete capability, such as reading email, executing shell commands,

agentsarxiv-cs-ai
5 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SUDP: Secret-Use Delegation Protocol for Agentic Systems

DGX agent

arXiv:2604.24920v2 Announce Type: replace-cross Abstract: Agentic systems increasingly act with user secrets for APIs, messaging platforms, and cloud services. Today's bearer-secret interfaces impleme

agentsarxiv-cs-ai
5 May 2026
Agents

Build, Judge, Optimize: A Blueprint for Continuous Improvement of Multi-Agent Consumer Assistants

DGX agent

arXiv:2603.03565v2 Announce Type: replace-cross Abstract: Conversational shopping assistants (CSAs) represent a compelling application of agentic AI, but moving from prototype to production reveals tw

agentsarxiv-cs-cl
4 May 2026
Agents

Crab: A Semantics-Aware Checkpoint/Restore Runtime for Agent Sandboxes

DGX agent

arXiv:2604.28138v1 Announce Type: cross Abstract: Autonomous agents act through sandboxed containers and microVMs whose state spans filesystems, processes, and runtime artifacts. Checkpoint and restor

agentsarxiv-cs-ai
1 May 2026
Agents

AGEL-Comp: A Neuro-Symbolic Framework for Compositional Generalization in Interactive Agents

DGX agent

arXiv:2604.26522v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents exhibit systemic failures in compositional generalization, limiting their robustness in interactive environments

agentsarxiv-cs-ai
30 Apr 2026
Model Releases

Exploring Reasoning Reward Model for Agents

DGX agent

arXiv:2601.22154v2 Announce Type: replace-cross Abstract: Agentic Reinforcement Learning (Agentic RL) has achieved notable success in enabling agents to perform complex reasoning and tool use. However

model-releasesarxiv-cs-cl
29 Apr 2026
Agents

DRACULA: Hunting for the Actions Users Want Deep Research Agents to Execute

DGX agent

arXiv:2604.23815v1 Announce Type: new Abstract: Scientific Deep Research (DR) agents answer user queries by synthesizing research papers into multi-section reports. User feedback can improve their uti

agentsarxiv-cs-cl
28 Apr 2026
Safety

Governing What You Cannot Observe: Adaptive Runtime Governance for Autonomous AI Agents

DGX agent

arXiv:2604.24686v1 Announce Type: new Abstract: Autonomous AI agents can remain fully authorized and still become unsafe as behavior drifts, adversaries adapt, and decision patterns shift without any

safetyarxiv-cs-ai
28 Apr 2026
Agents

The High Cost of Incivility: Quantifying Interaction Inefficiency via Multi-Agent Monte Carlo Simulations

DGX agent

arXiv:2512.08345v2 Announce Type: replace Abstract: Workplace toxicity is widely recognized as detrimental to organizational culture, yet quantifying its direct impact on operational efficiency remain

agentsarxiv-cs-ai
28 Apr 2026
Safety

Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems

DGX agent

arXiv:2604.22154v1 Announce Type: cross Abstract: Emerging AI systems in behavioral health and psychiatry use multi-step or multi-agent LLM pipelines for tasks like assessing self-harm risk and screen

safetyarxiv-cs-ai
27 Apr 2026
Agents

An Overlay Multicast Routing Method Based on Network Situational Awareness and Hierarchical Multi-Agent Reinforcement Learning

DGX agent

arXiv:2602.13211v2 Announce Type: replace-cross Abstract: Compared with IP multicast, Overlay Multicast (OM) offers better compatibility and flexible deployment in heterogeneous, cross-domain networks

agentsarxiv-cs-ai
24 Apr 2026
Agents

The Last Harness You'll Ever Build

DGX agent

arXiv:2604.21003v1 Announce Type: new Abstract: AI agents are increasingly deployed on complex, domain-specific workflows -- navigating enterprise web applications that require dozens of clicks and fo

agentsarxiv-cs-ai
24 Apr 2026
Hardware

A Delta-Aware Orchestration Framework for Scalable Multi-Agent Edge Computing

DGX agent

arXiv:2604.20129v1 Announce Type: new Abstract: The Synergistic Collapse occurs when scaling beyond 100 agents causes superlinear performance degradation that individual optimizations cannot prevent.

hardwarearxiv-cs-lg
23 Apr 2026
Local Ai

CEDAR: Context Engineering for Agentic Data Science

DGX agent

arXiv:2601.06606v2 Announce Type: replace-cross Abstract: We demonstrate CEDAR, an application for automating data science (DS) tasks with an agentic setup. Solving DS problems with LLMs is an underex

local-aiarxiv-cs-ai
23 Apr 2026
Agents

CHORUS: An Agentic Framework for Generating Realistic Deliberation Data

DGX agent

arXiv:2604.20651v1 Announce Type: new Abstract: Understanding the intricate dynamics of online discourse depends on large-scale deliberation data, a resource that remains scarce across interactive web

agentsarxiv-cs-ai
23 Apr 2026
Local Ai

Device-Native Autonomous Agents for Privacy-Preserving Negotiations

DGX agent

arXiv:2601.00911v3 Announce Type: replace-cross Abstract: Automated negotiations in insurance and business-to-business (B2B) commerce encounter substantial challenges. Current systems force a trade-of

local-aiarxiv-cs-ai
23 Apr 2026
Safety

FSFM: A Biologically-Inspired Framework for Selective Forgetting of Agent Memory

DGX agent

arXiv:2604.20300v1 Announce Type: new Abstract: For LLM agents, memory management critically impacts efficiency, quality, and security. While much research focuses on retention, selective forgetting--

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

Prism: An Evolutionary Memory Substrate for Multi-Agent Open-Ended Discovery

DGX agent

arXiv:2604.19795v1 Announce Type: new Abstract: We introduce prism{} (extbf{P}robabilistic extbf{R}etrieval with extbf{I}nformation-extbf{S}tratified extbf{M}emory), an evolutionary memory substrate f

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

AblateCell: A Reproduce-then-Ablate Agent for Virtual Cell Repositories

DGX agent

arXiv:2604.19606v1 Announce Type: new Abstract: Systematic ablations are essential to attribute performance gains in AI Virtual Cells, yet they are rarely performed because biological repositories are

agentsarxiv-cs-ai
22 Apr 2026
Agents

Advancing MAPF Toward the Real World: A Scalable Multi-Agent Realistic Testbed (SMART)

DGX agent

arXiv:2503.04798v3 Announce Type: replace Abstract: We present Scalable Multi-Agent Realistic Testbed (SMART), a realistic and efficient software tool for evaluating Multi-Agent Path Finding (MAPF) al

agentsarxiv-cs-ro
21 Apr 2026
Safety

DynaWeb: Model-Based Reinforcement Learning of Web Agents

DGX agent

arXiv:2601.22149v2 Announce Type: replace Abstract: The development of autonomous web agents, powered by Large Language Models (LLMs) and reinforcement learning (RL), represents a significant step tow

safetyarxiv-cs-cl
21 Apr 2026
Local Ai

GraSP: Graph-Structured Skill Compositions for LLM Agents

DGX agent

arXiv:2604.17870v1 Announce Type: new Abstract: Skill ecosystems for LLM agents have matured rapidly, yet recent benchmarks show that providing agents with more skills does not monotonically improve p

local-aiarxiv-cs-cl
21 Apr 2026
Model Releases

HiRAS: A Hierarchical Multi-Agent Framework for Paper-to-Code Generation and Execution

DGX agent

arXiv:2604.17745v1 Announce Type: new Abstract: Recent advances in large language models have highlighted their potential to automate computational research, particularly reproducing experimental resu

model-releasesarxiv-cs-cl
21 Apr 2026
Agents

LLM as a Tool, Not an Agent: Code-Mined Tree Transformations for Neural Architecture Search

DGX agent

arXiv:2604.16555v1 Announce Type: cross Abstract: Neural Architecture Search (NAS) aims to automatically discover high-performing deep neural network (DNN) architectures. However, conventional algorit

agentsarxiv-cs-cv
21 Apr 2026
Agents

Memory Centric Power Allocation for Multi-Agent Embodied Question Answering

DGX agent

arXiv:2604.17810v1 Announce Type: new Abstract: This paper considers multi-agent embodied question answering (MA-EQA), which aims to query robot teams on what they have seen over a long horizon. In co

agentsarxiv-cs-ro
21 Apr 2026
Agents

Self-Reasoning Agentic Framework for Narrative Product Grid-Collage Generation

DGX agent

arXiv:2604.16958v1 Announce Type: new Abstract: Narrative-driven product photography has become a prevalent paradigm in modern marketing, as coherent visual storytelling helps convey product value and

agentsarxiv-cs-cv
21 Apr 2026
Agents

FACTS: Table Summarization via Offline Template Generation with Agentic Workflows

DGX agent

arXiv:2510.13920v2 Announce Type: replace Abstract: Query-focused table summarization requires generating natural language summaries of tabular data conditioned on a user query, enabling users to acce

agentsarxiv-cs-cl
20 Apr 2026
Agents

The World Leaks the Future: Harness Evolution for Future Prediction Agents

DGX agent

arXiv:2604.15719v1 Announce Type: new Abstract: Many consequential decisions must be made before the relevant outcome is known. Such problems are commonly framed as future prediction, where an LLM age

agentsarxiv-cs-ai
20 Apr 2026
Hardware

Towards Understanding, Analyzing, and Optimizing Agentic AI Execution: A CPU-Centric Perspective

DGX agent

arXiv:2511.00739v3 Announce Type: replace Abstract: Agentic AI serving converts monolithic LLM-based inference to autonomous problem-solvers that can plan, call tools, perform reasoning, and adapt on

hardwarearxiv-cs-ai
20 Apr 2026
Model Releases

AgentGA: Evolving Code Solutions in Agent-Seed Space

DGX agent

arXiv:2604.14655v1 Announce Type: cross Abstract: We present AgentGA, a framework that evolves autonomous code-generation runs by optimizing the agent seed: the task prompt plus optional parent archiv

model-releasesarxiv-cs-lg
17 Apr 2026
Agents

APEX-MEM: Agentic Semi-Structured Memory with Temporal Reasoning for Long-Term Conversational AI

DGX agent

arXiv:2604.14362v1 Announce Type: new Abstract: Large language models still struggle with reliable long-term conversational memory: simply enlarging context windows or applying naive retrieval often i

agentsarxiv-cs-cl
17 Apr 2026
Safety

Towards Scalable Lightweight GUI Agents via Multi-role Orchestration

DGX agent

arXiv:2604.13488v1 Announce Type: new Abstract: Autonomous Graphical User Interface (GUI) agents powered by Multimodal Large Language Models (MLLMs) enable digital automation on end-user devices. Whil

safetyarxiv-cs-ai
17 Apr 2026
Safety

Your LLM Agents are Temporally Blind: The Misalignment Between Tool Use Decisions and Human Time Perception

DGX agent

arXiv:2510.23853v3 Announce Type: replace Abstract: Large language model (LLM) agents are increasingly used to interact with and execute tasks in dynamic environments. However, a critical yet overlook

safetyarxiv-cs-cl
17 Apr 2026
Agents

Agentic Conversational Search with Contextualized Reasoning via Reinforcement Learning

DGX agent

arXiv:2601.13115v2 Announce Type: replace Abstract: Large Language Models (LLMs) have become a popular interface for human-AI interaction, supporting information seeking and task assistance through na

agentsarxiv-cs-cl
16 Apr 2026
Safety

Beyond Arrow's Impossibility: Fairness as an Emergent Property of Multi-Agent Collaboration

DGX agent

arXiv:2604.13705v1 Announce Type: new Abstract: Fairness in language models is typically studied as a property of a single, centrally optimized model. As large language models become increasingly agen

safetyarxiv-cs-cl
16 Apr 2026
Model Releases

Bridging MARL to SARL: An Order-Independent Multi-Agent Transformer via Latent Consensus

DGX agent

arXiv:2604.13472v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning (MARL) is widely used to address large joint observation and action spaces by decomposing a centralized c

model-releasesarxiv-cs-lg
16 Apr 2026
Agents

LEO-RobotAgent: A General-purpose Robotic Agent for Language-driven Embodied Operator

DGX agent

arXiv:2512.10605v2 Announce Type: replace Abstract: We propose LEO-RobotAgent, a general-purpose language-driven intelligent agent framework for robots. Under this framework, LLMs can operate differen

agentsarxiv-cs-ro
16 Apr 2026
Agents

ALL-FEM: Agentic Large Language models Fine-tuned for Finite Element Methods

DGX agent

arXiv:2603.21011v2 Announce Type: replace-cross Abstract: Finite element (FE) analysis guides the design and verification of nearly all manufactured objects. It is at the core of computational enginee

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

CODESTRUCT: Code Agents over Structured Action Spaces

DGX agent

arXiv:2604.05407v2 Announce Type: replace Abstract: LLM-based code agents treat repositories as unstructured text, applying edits through brittle string matching that frequently fails due to formattin

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

SecureWebArena: A Holistic Security Evaluation Benchmark for LVLM-based Web Agents

DGX agent

arXiv:2510.10073v2 Announce Type: replace-cross Abstract: Large vision-language model (LVLM)-based web agents are emerging as powerful tools for automating complex online tasks. However, when deployed

model-releasesarxiv-cs-cv
15 Apr 2026
Agents

Collaborative Multi-Agent Scripts Generation for Enhancing Imperfect-Information Reasoning in Murder Mystery Games

DGX agent

arXiv:2604.11741v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown impressive capabilities in perceptual tasks, yet they degrade in complex multi-hop reasoning under multiplayer

agentsarxiv-cs-ai
14 Apr 2026
Safety

Decision-Theoretic Safety Assessment of Persona-Driven Multi-Agent Systems in O-RAN

DGX agent

arXiv:2604.09682v1 Announce Type: cross Abstract: Autonomous network management in Open Radio Access Networks requires intelligent decision making across conflicting objectives, yet existing LLM based

safetyarxiv-cs-ai
14 Apr 2026
Agents

HTAA: Enhancing LLM Planning via Hybrid Toolset Agentization & Adaptation

DGX agent

arXiv:2604.10917v1 Announce Type: new Abstract: Enabling large language models to scale and reliably use hundreds of tools is critical for real-world applications, yet challenging due to the inefficie

agentsarxiv-cs-cl
14 Apr 2026
Safety

MobiFlow: Real-World Mobile Agent Benchmarking through Trajectory Fusion

DGX agent

arXiv:2604.09587v1 Announce Type: new Abstract: Mobile agents can autonomously complete user-assigned tasks through GUI interactions. However, existing mainstream evaluation benchmarks, such as Androi

safetyarxiv-cs-ai
14 Apr 2026
Agents

WaterAdmin: Orchestrating Community Water Distribution Optimization via AI Agents

DGX agent

arXiv:2604.10343v1 Announce Type: new Abstract: We study the operation of community water systems, where pumps and valves must be scheduled to reliably meet water demands while minimizing energy consu

agentsarxiv-cs-lg
14 Apr 2026
Agents

H-AdminSim: A Multi-Agent Simulator for Realistic Hospital Administrative Workflows with FHIR Integration

DGX agent

arXiv:2602.05407v2 Announce Type: replace Abstract: Hospital administration departments handle a wide range of operational tasks and, in large hospitals, process over 10,000 requests per day, driving

agentsarxiv-cs-ai
13 Apr 2026
Model Releases

Many-Tier Instruction Hierarchy in LLM Agents

DGX agent

arXiv:2604.09443v1 Announce Type: cross Abstract: Large language model agents receive instructions from many sources-system messages, user prompts, tool outputs, and more-each carrying different level

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

ReplicatorBench: Benchmarking LLM Agents for Replicability in Social and Behavioral Sciences

DGX agent

arXiv:2602.11354v2 Announce Type: replace Abstract: The literature has witnessed an emerging interest in AI agents for automated assessment of scientific papers. Existing benchmarks focus primarily on

model-releasesarxiv-cs-ai
13 Apr 2026
← Previous
1…2324252627…233
Next →