AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Model Releases

STARS: Skill-Triggered Audit for Request-Conditioned Invocation Safety in Agent Systems

DGX agent

arXiv:2604.10286v1 Announce Type: new Abstract: Autonomous language-model agents increasingly rely on installable skills and tools to complete user tasks. Static skill auditing can expose capability s

model-releasesarxiv-cs-ai
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

SWE-AGILE: A Software Agent Framework for Efficiently Managing Dynamic Reasoning Context

DGX agent

arXiv:2604.11716v1 Announce Type: new Abstract: Prior representative ReAct-style approaches in autonomous Software Engineering (SWE) typically lack the explicit System-2 reasoning required for deep an

agentsarxiv-cs-ai
14 Apr 2026
Agents

CodeScout: Contextual Problem Statement Enhancement for Software Agents

DGX agent

arXiv:2603.05744v2 Announce Type: replace Abstract: Current AI-powered code assistance tools often struggle with poorly-defined problem statements that lack sufficient task context and requirements sp

agentsarxiv-cs-cl
13 Apr 2026
Agents

Gen-n-Val: Agentic Image Data Generation and Validation

DGX agent

arXiv:2506.04676v2 Announce Type: replace-cross Abstract: The data scarcity, label noise, and long-tailed category imbalance remain important and unresolved challenges in many computer vision tasks, s

agentsarxiv-cs-ai
13 Apr 2026
Agents

Sustained Impact of Agentic Personalisation in Marketing: A Longitudinal Case Study

DGX agent

arXiv:2604.08621v1 Announce Type: new Abstract: In consumer applications, Customer Relationship Management (CRM) has traditionally relied on the manual optimisation of static, rule-based messaging str

agentsarxiv-cs-ai
13 Apr 2026
Safety

Wireless Communication Enhanced Value Decomposition for Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.08728v1 Announce Type: new Abstract: Cooperation in multi-agent reinforcement learning (MARL) benefits from inter-agent communication, yet most approaches assume idealized channels and exis

safetyarxiv-cs-lg
13 Apr 2026
Model Releases

ACIArena: Toward Unified Evaluation for Agent Cascading Injection

DGX agent

arXiv:2604.07775v1 Announce Type: cross Abstract: Collaboration and information sharing empower Multi-Agent Systems (MAS) but also introduce a critical security risk known as Agent Cascading Injection

model-releasesarxiv-cs-cl
10 Apr 2026
Safety

AgentCity: Constitutional Governance for Autonomous Agent Economies via Separation of Power

DGX agent

arXiv:2604.07007v1 Announce Type: cross Abstract: Autonomous AI agents are beginning to operate across organizational boundaries on the open internet -- discovering, transacting with, and delegating t

safetyarxiv-cs-ai
10 Apr 2026
Agents

ClawLess: A Security Model of AI Agents

DGX agent

arXiv:2604.06284v1 Announce Type: cross Abstract: Autonomous AI agents powered by Large Language Models can reason, plan, and execute complex tasks, but their ability to autonomously retrieve informat

agentsarxiv-cs-ai
10 Apr 2026
Agents

Density-Driven Optimal Control: Convergence Guarantees for Stochastic LTI Multi-Agent Systems

DGX agent

arXiv:2604.08495v1 Announce Type: cross Abstract: This paper addresses the decentralized non-uniform area coverage problem for multi-agent systems, a critical task in missions with high spatial priori

agentsarxiv-cs-ro
10 Apr 2026
Agents

Logical Robots: Declarative Multi-Agent Programming in Logica

DGX agent

arXiv:2604.06629v1 Announce Type: cross Abstract: We present Logical Robots, an interactive multi-agent simulation platform where autonomous robot behavior is specified declaratively in the logic prog

agentsarxiv-cs-ai
10 Apr 2026
Agents

ReDAct: Uncertainty-Aware Deferral for LLM Agents

DGX agent

arXiv:2604.07036v1 Announce Type: cross Abstract: Recently, LLM-based agents have become increasingly popular across many applications, including complex sequential decision-making problems. However,

agentsarxiv-cs-lg
10 Apr 2026
Agents

Stay Focused: Problem Drift in Multi-Agent Debate

DGX agent

arXiv:2502.19559v3 Announce Type: replace Abstract: Multi-agent debate - multiple instances of large language models discussing problems in turn-based interaction - has shown promise for solving knowl

agentsarxiv-cs-cl
10 Apr 2026
Agents

How to Dogfood Your AI Chat Agent: A Three-Layer Evaluation Framework with Goal-Directed NPC Simulation

DGX agent

arXiv:2608.09939v1 Announce Type: cross Abstract: Production teams deploying LLM chat agents face a specific quality assurance gap: existing evaluation tools test individual responses or simulate soci

agentsarxiv-cs-ai
12 Aug 2026
Agents

OpenPM: Auditable Point-in-Time Evaluation for LLM Portfolio-Management Agents

DGX agent

arXiv:2608.09988v1 Announce Type: cross Abstract: Large language models are increasingly used to read markets, assess risk, and allocate capital. However, reported results for LLM trading agents can b

agentsarxiv-cs-cl
12 Aug 2026
Agents

The Signal Rail: A Deterministic Motion Grammar for Communicating Conversational Agent State in Terminal Interfaces

DGX agent

arXiv:2608.10689v1 Announce Type: cross Abstract: Terminal interfaces to conversational agents report rich internal state (listening, thinking, executing tools, awaiting input, failing) almost entirel

agentsarxiv-cs-cl
12 Aug 2026
Agents

From Relevance to Execution Utility: Reward-Aware Dynamic Execution Gating for Skill-Based LLM Agents

DGX agent

arXiv:2608.09168v1 Announce Type: new Abstract: Agent skills are increasingly used to equip large language model (LLM) agents with reusable procedural knowledge. Although recent work has substantially

agentsarxiv-cs-ai
11 Aug 2026
Agents

Lingjing: A Simulation Testbed for Multi-Agent Embodied Tasks in Open-Ended Cities

DGX agent

arXiv:2608.08045v1 Announce Type: new Abstract: Urban embodied intelligence requires coordination among heterogeneous agents (e.g., UAVs, ground robots, and autonomous vehicles) in dynamic cities. Sim

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

OmnilingualGAIA2: Evaluating the Multilingual Gap in Frontier AI Agents

DGX agent

arXiv:2608.08775v1 Announce Type: new Abstract: Agentic benchmarks aim to measure how well AI agents plan, search, execute, and recover within realistic multi-tool environments, but they are almost ex

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

Deal Me Maybe: The Role of Emotions in Multi-Agent Negotiation

DGX agent

arXiv:2608.06922v1 Announce Type: new Abstract: Negotiation is a demanding social task for LLM agents, requiring strategic reasoning, persuasion, and interpersonal adaptation. Yet existing benchmarks

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

NiyamAI - An Intent-Bound AI Agent with Cryptographically Verifiable Guardrails using Zero-Knowledge Proofs

DGX agent

arXiv:2608.07167v1 Announce Type: new Abstract: Giving an AI agent the ability to send emails, query databases, or execute commands is useful--until the agent is tricked into doing something it should

model-releasesarxiv-cs-ai
10 Aug 2026
Agents

Scalable Long-Horizon Planning with Staggered Updates for Lifelong MAPF

DGX agent

arXiv:2608.06702v1 Announce Type: cross Abstract: Lifelong Multi-Agent Path Finding (LMAPF) requires generating collision-free paths for large agent fleets under strict real-time constraints. Reactive

agentsarxiv-cs-ai
10 Aug 2026
Safety

The Optimizer Is the Agent: Reasoning-Driven Search across Prompts, Programs, and ML Workflows

DGX agent

arXiv:2608.06714v1 Announce Type: new Abstract: Recent systems for optimizing prompts, programs, and ML workflows typically rely on explicit outer-loop controllers such as evolutionary search, bandits

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

Unified Agent: Managing Interactions across Devices

DGX agent

arXiv:2608.05729v1 Announce Type: new Abstract: As capabilities rapidly increase, AI agents can move from running inside one app to acting across a user's devices over time. Yet existing agent systems

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

Learning Sexism Detection Using Multi-Agent Perspectivist Preference Optimization

DGX agent

arXiv:2608.04056v1 Announce Type: new Abstract: When people label text for sexism, they often disagree, and not because some of them are wrong: they genuinely perceive sexism differently. Most NLP sys

agentsarxiv-cs-cl
6 Aug 2026
Agents

Terminal Agents Suffice for Enterprise Automation

DGX agent

arXiv:2604.00073v3 Announce Type: replace-cross Abstract: There has been growing interest in building agents that can interact with digital platforms to execute meaningful enterprise tasks autonomousl

agentsarxiv-cs-ai
6 Aug 2026
Agents

Formal Verification of Agentic Systems over Operational Data

DGX agent

arXiv:2608.03609v1 Announce Type: new Abstract: Agentic systems driven by large language models (LLMs) are increasingly deployed in real-world workflows where they act on persistent operational data.

agentsarxiv-cs-ai
5 Aug 2026
Agents

Internalising the Identity Primitive: Cryptographic Individuality for an Autonomous Agent on a Public Blockchain

DGX agent

arXiv:2608.02986v1 Announce Type: cross Abstract: A software agent on a public blockchain accumulates authority and economic stakes, raising the engineering question of what makes it count as an indiv

agentsarxiv-cs-ai
5 Aug 2026
Agents

MemoryForge: Synthesize Lifelong Memory for Human-Like LLM Agents

DGX agent

arXiv:2608.00007v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with human-like personas is crucial for agentic applications, such as role-play and user simulation. Traditional

agentsarxiv-cs-cl
4 Aug 2026
Agents

OTAP: Structure-Aware Optimal Transport for Evaluating Planning and Execution in Agent Trajectories

DGX agent

arXiv:2607.17082v2 Announce Type: replace-cross Abstract: Large language model agents solve tasks by generating trajectories that interleave planning, tool calls, and intermediate results. Current eva

agentsarxiv-cs-cl
4 Aug 2026
Safety

Training Small LLMs as Spatial Multi-Agent Policies

DGX agent

arXiv:2608.01425v1 Announce Type: cross Abstract: Training LLM-based multi-agent systems with multi-agent reinforcement learning is rapidly gaining traction, and a parallel line of work argues that su

safetyarxiv-cs-lg
4 Aug 2026
Model Releases

Model or Harness? An Interaction-Centric Taxonomy for Localizing Agent Failures

DGX agent

arXiv:2607.28802v1 Announce Type: new Abstract: Existing evaluations often reduce agent failures to system-level outcomes, obscuring where the fault originated and which intervention would improve the

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Safety, or Just Capability? A Validity Audit of Agent-Safety Benchmarks

DGX agent

arXiv:2607.28685v1 Announce Type: new Abstract: Agent-safety benchmarks measure different behaviors, and their scores get quoted interchangeably as an agent's safety. We treat four of them (R-Judge, I

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

A First Look at Coding Agents' Compliance with AI Contribution Rules in Open-Source Communities

DGX agent

arXiv:2607.26819v1 Announce Type: cross Abstract: Open source communities have been flooded with AI-generated contributions. In defense, they have written contribution rules to regulate coding agents'

agentsarxiv-cs-ai
31 Jul 2026
Agents

Beacon: Knowing When and How to Perform Agentic Visual Reasoning

DGX agent

arXiv:2607.28595v1 Announce Type: new Abstract: The fundamental goal of agentic visual reasoning is to improve the success rate of multimodal large language models (MLLMs) on complex tasks, rather tha

agentsarxiv-cs-cv
31 Jul 2026
Safety

Eco3S: Complex Socio-Economic System Simulation via Agent-Based Models

DGX agent

arXiv:2607.26588v1 Announce Type: new Abstract: The rapid development of large language models (LLMs) has renewed interest in agent-based modeling (ABM). However, current LLM-based ABM research faces

safetyarxiv-cs-ai
31 Jul 2026
Safety

SimpleWikiSearch: A Clean Offline Wikipedia Environment for Agentic Search

DGX agent

arXiv:2607.26070v1 Announce Type: cross Abstract: Large language model (LLM)-based agentic search systems are often evaluated as if the underlying LLM were the only component that matters, yet their m

safetyarxiv-cs-ai
31 Jul 2026
Agents

SpecFirst: Behavioral Specification Elicitation as a First-Class Step in Agent-Based Program Synthesis from Scratch

DGX agent

arXiv:2607.27167v1 Announce Type: cross Abstract: LLM-based agents excel at software engineering tasks where an existing codebase provides context, but constructing a program from scratch remains fund

agentsarxiv-cs-cl
30 Jul 2026
Agents

Agentic Autoresearch for CT Reconstruction

DGX agent

arXiv:2607.22824v1 Announce Type: cross Abstract: Comparing CT reconstruction methods fairly is labor-intensive and largely manual, and many benchmarks use idealized data. We ask whether a large langu

agentsarxiv-cs-ai
28 Jul 2026
Safety

ARdena: Scenario-driven control of real-time LLM agents

DGX agent

arXiv:2607.22651v1 Announce Type: new Abstract: Large language models (LLMs) have enabled increasingly capable conversational agents, but reliably controlling their behavior in real-time interactive e

safetyarxiv-cs-ai
28 Jul 2026
Agents

Commitment To Cooperation With Self-Negotiated Contracts

DGX agent

arXiv:2607.22750v1 Announce Type: new Abstract: As AI agents operate with increasing autonomy in a multi-agent world, they will need to learn to cooperate with other agents and with humans to generate

agentsarxiv-cs-ai
28 Jul 2026
Agents

From Cognitive Architectures to Language Agents: A Mechanism-Level Review of Lineage, Convergence, and Migration Gaps

DGX agent

arXiv:2607.23942v1 Announce Type: new Abstract: Memory, planning, reflection, and tool use are often compared as feature labels, obscuring the control semantics that determine how an agent actually ru

agentsarxiv-cs-ai
28 Jul 2026
Agents

SCTA: An Agentic Framework for Stable and Interpretable Target Gene Discovery from Single-Cell RNA Sequencing

DGX agent

arXiv:2607.23821v1 Announce Type: new Abstract: Identifying therapeutic target genes from single-cell RNA sequencing (scRNA-seq) data remains a fundamental challenge in translational biology. Unlike b

agentsarxiv-cs-lg
28 Jul 2026
Agents

SeekJudge: A Practical Reward Framework for Reinforcement Learning in Computer-Use Agents

DGX agent

arXiv:2607.23263v1 Announce Type: new Abstract: Deciding whether a trajectory actually fulfills its instruction governs how we measure computer-use agents on long-horizon graphical-user-interface task

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Who Pays the Price? Stakeholder-Centric Prompt Injection Benchmarking for Real-world Web Agents

DGX agent

arXiv:2606.13385v2 Announce Type: replace-cross Abstract: LLM-based web agents are increasingly deployed in real-world settings such as e-commerce, where they interact extensively with untrusted web c

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Can AI Agents Really Complete RTL-to-GDS? Lessons from Benchmarking Tool-Interactive EDA Workflows

DGX agent

arXiv:2607.17528v3 Announce Type: replace Abstract: Large language model (LLM) agents are extending electronic design automation (EDA) beyond static RTL generation toward long-horizon, tool-interactiv

agentsarxiv-cs-ai
24 Jul 2026
Agents

Causal-AgentIR: Self-Evolving Causal Memory for Adaptive Image Restoration Agents

DGX agent

arXiv:2607.21125v1 Announce Type: new Abstract: Image restoration agents have recently emerged as a flexible paradigm for handling diverse and unpredictable degradations in real-world scenarios. Exist

agentsarxiv-cs-cv
24 Jul 2026
Model Releases

WebCoach: Self-Evolving Web Agents with Cross-Session Memory Guidance

DGX agent

arXiv:2511.12997v2 Announce Type: replace Abstract: Multimodal LLM-powered agents have recently demonstrated impressive capabilities in web navigation, enabling agents to complete complex browsing tas

model-releasesarxiv-cs-ai
24 Jul 2026
← Previous
1…1617181920…230
Next →