AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Agents

CORVUS: Context Optimization and Reduction Via Underlying Synchronization for LLM Coding Agents

DGX agent

arXiv:2607.22711v1 Announce Type: cross Abstract: LLM coding agents operate by constructing trajectories that accumulate reasoning, tool calls, and results to enable multi-step decision-making. Howeve

agentsarxiv-cs-ai
28 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Decentralized Granular Access Control for Agentic AI Systems in Critical Infrastructure

DGX agent

arXiv:2607.22611v1 Announce Type: new Abstract: The deployment of autonomous AI agents in production infrastructure introduces fundamental security challenges that traditional role-based access contro

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

False Prophets: On the Security of World Models in Agentic Systems

DGX agent

arXiv:2607.23147v1 Announce Type: cross Abstract: Large language models now power autonomous agents capable of complex, multi-step tasks in different environments. Accurate and reliable execution of t

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Hybrid Advantage Estimation with Unified Critic for VLM Agentic Reinforcement Learning

DGX agent

arXiv:2607.23605v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) now act as agents in interactive environments, where success requires coherent reasoning and decision-making across

agentsarxiv-cs-ai
28 Jul 2026
Agents

PD^3: A Project Duplication Detection Framework via Adapted Multi-Agent Debate

DGX agent

arXiv:2505.17492v2 Announce Type: replace Abstract: Project duplication detection is critical for project quality assessment because it helps avoid investment in repeated proposals. Existing methods u

agentsarxiv-cs-ai
28 Jul 2026
Agents

Scaling GUI Agents with Visual State Transitions

DGX agent

arXiv:2607.24112v1 Announce Type: new Abstract: We introduce State Transition Pretraining (STP) as a new scaling axis for GUI agents. During the STP stage, we continually pretrain a unified multimodal

agentsarxiv-cs-ai
28 Jul 2026
Safety

Self-Authored Verification Is Unreliable in Heuristic Self-Improving Agents

DGX agent

arXiv:2607.24300v1 Announce Type: new Abstract: Self-improving agents accumulate capability by repeatedly rewriting procedural policies, controllers, or heuristic rules. They typically rely on self-au

safetyarxiv-cs-cl
28 Jul 2026
Agents

Separating Capability from Permission: A Governance Framework for Agentic AI Autonomy Levels

DGX agent

arXiv:2607.23438v1 Announce Type: new Abstract: As AI systems increasingly exhibit agentic behavior, discussions of autonomy often conflate what systems are technically capable of doing with what they

agentsarxiv-cs-ai
28 Jul 2026
Agents

SpecBox: Speculative Sandbox Scheduling for Efficient LLM Agent Serving

DGX agent

arXiv:2607.23933v1 Announce Type: cross Abstract: As LLM agents increasingly rely on the Model Context Protocol (MCP) to invoke isolated external sandboxes, disaggregated sandbox deployment introduces

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Success Is Not Self-Explanatory: Auditing Success Provenance in Agent Evaluation

DGX agent

arXiv:2607.24054v1 Announce Type: new Abstract: A correct answer can conceal why an agent succeeded. Once agents change their information state during evaluation, correctness no longer distinguishes i

model-releasesarxiv-cs-ai
28 Jul 2026
Model Releases

Agentic Designer: Progressive Multi-Agent Collaboration for Structure-Aware Interior Layout Generation

DGX agent

arXiv:2607.20866v1 Announce Type: new Abstract: Generating realistic interior furniture layouts that strictly adhere to architectural constraints (e.g., walls, doors, and windows) remains a fundamenta

model-releasesarxiv-cs-cv
24 Jul 2026
Model Releases

IssueTrojanBench: Benchmarking AI Coding Agents Against Malicious Issue Requests

DGX agent

arXiv:2607.20759v1 Announce Type: cross Abstract: AI coding agents powered by LLMs are increasingly integrated into real-world software development, where they generate, edit, and execute code with au

model-releasesarxiv-cs-ai
24 Jul 2026
Safety

Agentic Real2Sim: Physics-based World Modeling with Vision-Language Agents

DGX agent

arXiv:2607.19190v2 Announce Type: replace-cross Abstract: Real-to-sim conversion for robotic interaction with objects remains labor-intensive because it requires more than visual reconstruction: a str

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

CAVA: Canonical Action Verification and Attestation for Runtime Governance of Agentic AI Systems

DGX agent

arXiv:2607.13716v1 Announce Type: new Abstract: Agentic AI systems increasingly act through heterogeneous runtimes: local coding hooks, SDK tools, browser automation, managed-agent traces, API gateway

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Self-Improvements in Modern Agentic Systems: A Survey

DGX agent

arXiv:2607.13104v1 Announce Type: new Abstract: Self-improving autonomous agents are moving from research prototypes to deployed systems. The primary goal is controllable evolution, or adaptation, fro

agentsarxiv-cs-ai
16 Jul 2026
Agents

Self-Improving AI Coding Agents Through Accumulated Behavioral Rules: A Closed-Loop Framework

DGX agent

arXiv:2607.13091v1 Announce Type: cross Abstract: LLM-based coding agents repeat the same classes of mistakes across sessions because they lack a mechanism to retain corrections from human review feed

agentsarxiv-cs-ai
16 Jul 2026
Model Releases

When Agents Disagree With Themselves: Behavioral Consistency as an Uncertainty Signal for LLM Agents

DGX agent

arXiv:2602.11619v2 Announce Type: replace Abstract: Running the same LLM agent on identical inputs yields 2.3-4.2 distinct action sequences per 10 runs; this behavioral variance constitutes a training

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Speculate with Memory: Lossless Acceleration for LLM Agents

DGX agent

arXiv:2607.12236v1 Announce Type: cross Abstract: Speculative execution accelerates LLM agents by using a smaller, cheaper model to predict and pre-launch the next step while the environment is idle.

agentsarxiv-cs-cl
15 Jul 2026
Model Releases

Track, Rank, Crack: Epistemic Working Memory Scales Multi-Hop Reasoning in Language Agents

DGX agent

arXiv:2607.12267v1 Announce Type: cross Abstract: Language agents that interleave reasoning and tool use degrade sharply as reasoning chains lengthen, even when each individual step is easy. We trace

model-releasesarxiv-cs-ai
15 Jul 2026
Agents

Effective Strategies for Asynchronous Software Engineering Agents

DGX agent

arXiv:2603.21489v2 Announce Type: replace-cross Abstract: AI agents have become increasingly capable at isolated software engineering (SWE) tasks such as resolving issues on Github. Yet long-horizon t

agentsarxiv-cs-ai
9 Jul 2026
Agents

From Atomic Actions to Standard Operating Procedures: Iterative Tool Optimization for Self-Evolving LLM Agents

DGX agent

arXiv:2607.07321v1 Announce Type: new Abstract: Tool utilization enables Large Language Model (LLM) agents to interact with the real world and resolve complex tasks. However, existing agent frameworks

agentsarxiv-cs-ai
9 Jul 2026
Agents

AppAgent: Multimodal Agents as Smartphone Users

DGX agent

arXiv:2312.13771v3 Announce Type: replace Abstract: Recent advancements in large language models (LLMs) have led to the creation of intelligent agents capable of performing complex tasks. This paper i

agentsarxiv-cs-cv
7 Jul 2026
Safety

Autonomous Information Seeking: A Roadmap for Agentic Recommender Systems

DGX agent

arXiv:2607.04433v1 Announce Type: cross Abstract: The rapid integration of large language model-based agents into recommender systems has driven a shift from static, ranking-based pipelines toward aut

safetyarxiv-cs-cl
7 Jul 2026
Agents

Is Agentic Code Review Helpful? Mining Developers' Feedback to CodeRabbit Reviews in the Wild

DGX agent

arXiv:2607.03316v1 Announce Type: cross Abstract: Agentic code review, where autonomous agents provide code review comments on pull requests, is increasingly integrated into development workflows, yet

agentsarxiv-cs-ai
7 Jul 2026
Agents

Web-CogReasoner: Towards Multimodal Knowledge-Induced Cognitive Reasoning for Web Agents

DGX agent

arXiv:2508.01858v3 Announce Type: replace-cross Abstract: Multimodal large-scale models have significantly advanced the development of web agents, enabling perception and interaction with digital envi

agentsarxiv-cs-ai
7 Jul 2026
Safety

Managed Autonomy at Runtime: Gear-Based Safety and Governance for Single- and Multi-Agent Cyber-Physical Systems

DGX agent

arXiv:2607.00334v1 Announce Type: new Abstract: Autonomous agents, whether LLM-driven software agents or robotic physical agents, face a common class of failure modes when operating without continuous

safetyarxiv-cs-ai
2 Jul 2026
Agents

RESOURCE2SKILL: Distilling Executable Agent Skills from Human-Created Multimodal Resources

DGX agent

arXiv:2606.29538v1 Announce Type: cross Abstract: Skills are a useful abstraction for software agents, turning human and agent experience into reusable procedural knowledge. Yet existing skill librari

agentsarxiv-cs-ai
30 Jun 2026
Safety

Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems

DGX agent

arXiv:2606.28425v1 Announce Type: cross Abstract: Increasingly autonomous agentic AI systems pose novel multi-agent risks, such as secret collusion via covert communication channels. The natural defen

safetyarxiv-cs-ai
30 Jun 2026
Safety

Agent-Native Immune System: Architecture, Taxonomy, and Engineering

DGX agent

arXiv:2606.28270v1 Announce Type: new Abstract: The transition from static chat bots to autonomous agents--equipped with persistent memory, tool-use protocols, and multi-agent collaboration--has funda

safetyarxiv-cs-ai
29 Jun 2026
Agents

A Deterministic Control Plane for LLM Coding Agents

DGX agent

arXiv:2606.26924v1 Announce Type: cross Abstract: LLM coding harnesses grant agents broad file and shell access, yet the configuration layer that steers them -- rules files, agent definitions, IDE-spe

agentsarxiv-cs-ai
26 Jun 2026
Agents

Agentic Collaborative Cognition for Zero-Shot 3D Understanding

DGX agent

arXiv:2606.24649v1 Announce Type: new Abstract: Recent advancements have explored agentic zero-shot 3D understanding by reformulating it as video keyframe understanding with Multimodal Large Language

agentsarxiv-cs-cv
24 Jun 2026
Tutorials

Weak-Driven Learning: How Weak Agents make Strong Agents Stronger

DGX agent

arXiv:2602.08222v2 Announce Type: replace Abstract: As post-training optimization becomes central to improving large language models, we observe a persistent saturation bottleneck: once models grow hi

tutorialsarxiv-cs-ai
9 Jun 2026
Agents

From Privacy to Workflow Integrity: Communication-Graph Metadata in Autonomous Agent Interoperability

DGX agent

arXiv:2606.07150v1 Announce Type: cross Abstract: Agent-interoperability protocols such as A2A and MCP standardize what agents say to one another, but assume address-based transport over HTTP(S). Such

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

ADK Arena: Evaluating Agent Development Kits via LLM-as-a-Developer

DGX agent

arXiv:2606.05548v1 Announce Type: cross Abstract: The rapid proliferation of Agent Development Kits (ADKs), SDK-level frameworks for building LLM-powered autonomous agents, has outpaced any empirical

model-releasesarxiv-cs-ai
6 Jun 2026
Agents

Ask or Assume? Uncertainty-Aware Clarification-Seeking in Coding Agents

DGX agent

arXiv:2603.26233v2 Announce Type: replace Abstract: As Large Language Model (LLM) agents are increasingly deployed in open-ended domains like software engineering, they frequently encounter underspeci

agentsarxiv-cs-cl
5 Jun 2026
Safety

Beyond Alignment: Value Diversity as a Collective Property in Multicultural Agent Systems

DGX agent

arXiv:2606.05985v1 Announce Type: new Abstract: Multicultural multi-agent systems are increasingly deployed in globally diverse settings, where different agents are grounded in different cultural back

safetyarxiv-cs-cl
5 Jun 2026
Agents

Entropy-Based Evaluation of AI Agents: A Lightweight Framework for Measuring Behavioral Patterns

DGX agent

arXiv:2606.05872v1 Announce Type: cross Abstract: AI agents are commonly evaluated using task success, reward, latency, and cost. These metrics are useful, but they often miss important aspects of age

agentsarxiv-cs-cv
5 Jun 2026
Agents

ShareVerse: Multi-Agent Consistent Video Generation for Shared World Modeling

DGX agent

arXiv:2603.02697v2 Announce Type: replace-cross Abstract: This paper presents ShareVerse, a video generation framework enabling multi-agent shared world modeling, addressing the gap in existing works

agentsarxiv-cs-ai
4 Jun 2026
Local Ai

Temporal Order Matters for Agentic Memory: Segment Trees for Long-Horizon Agents

DGX agent

arXiv:2606.04555v1 Announce Type: cross Abstract: Long-horizon conversational agents need to interact with users through evolving events, tasks, and goals. Such histories are naturally temporal, yet m

local-aiarxiv-cs-ai
4 Jun 2026
Agents

Economy of Minds: Emerging Multi-Agent Intelligence with Economic Interactions

DGX agent

arXiv:2606.02859v1 Announce Type: cross Abstract: How can a population of agents self-orchestrate and self-adapt into stronger collective intelligence without centralized control? Inspired by Friedric

agentsarxiv-cs-ai
3 Jun 2026
Agents

Overlaying Governance: A Compositional Authorization Framework for Delegation and Scope in Agentic AI

DGX agent

arXiv:2606.03518v1 Announce Type: new Abstract: As AI systems evolve from passive models into autonomous active agents capable of initiating actions, collaborating, and delegating tasks, the tradition

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

What Benchmarks Don't Measure: The Case for Evaluating Abstention Competence in Autonomous Agents

DGX agent

arXiv:2606.02965v1 Announce Type: new Abstract: Benchmarks for autonomous agents measure whether agents complete tasks, yet this framing is systematically blind to whether an agent should have proceed

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

FALAT: Tracing Failures in LLM Agent Trajectories via Dependency-Guided Search

DGX agent

arXiv:2606.00765v1 Announce Type: new Abstract: LLM-based agents increasingly solve complex tasks through long trajectories involving reasoning steps, tool calls, and inter-agent communication. Howeve

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Multi-Agent Computer Use

DGX agent

arXiv:2606.01533v1 Announce Type: cross Abstract: Computer use agents (CUAs) today are primarily deployed as single serial agents. This setup is suboptimal for complex long-horizon tasks that benefit

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Scaling Behavior of Single LLM-Driven Multi-Agent Systems

DGX agent

arXiv:2606.00655v1 Announce Type: cross Abstract: The burgeoning field of LLM-based Multi-Agent Systems (MAS) promises to tackle complex tasks through collaborative intelligence, yet fundamental quest

agentsarxiv-cs-ai
2 Jun 2026
Agents

Evolve as a Team: Collaborative Self-Evolution for LLM-based Multi-Agent Systems

DGX agent

arXiv:2605.29790v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) have emerged as an effective paradigm for complex and long-horizon tasks. However, in real-world tasks, MAS often

agentsarxiv-cs-ai
29 May 2026
Safety

ValueFlow: Measuring the Propagation of Value Perturbations in Multi-Agent LLM Systems

DGX agent

arXiv:2602.08567v2 Announce Type: replace-cross Abstract: Multi-agent large language model (LLM) systems increasingly consist of agents that observe and respond to one another's outputs. While value a

safetyarxiv-cs-cl
29 May 2026
Model Releases

AgentSociety: Incentivizing Agentic Social Intelligence

DGX agent

arXiv:2605.26203v1 Announce Type: cross Abstract: The success of deployed agents relies on their ability to handle open-ended user requests using their inherent capabilities, not only in solving reque

model-releasesarxiv-cs-ai
27 May 2026
← Previous
1…1415161718…230
Next →