AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Model Releases

StepJack: Benchmarking Computer-Use Agent Safety Against Multi-Step Indirect Prompt Injection

DGX agent

arXiv:2608.06477v1 Announce Type: cross Abstract: Computer-use agents (CUAs) face a growing threat from indirect prompt injection, where adversarial instructions are planted in the environment such as

model-releasesarxiv-cs-ai
10 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

EASy: Towards Efficient LLM-Based Agentic System

DGX agent

arXiv:2608.04588v1 Announce Type: cross Abstract: Agentic systems have emerged as a promising paradigm for solving complex tasks by coordinating specialized LLM-based agents. However, most existing sy

agentsarxiv-cs-ai
6 Aug 2026
Agents

MetaVideoAgent: Automated Video-Agent Evolution for Long-Form Video Understanding

DGX agent

arXiv:2608.04587v1 Announce Type: new Abstract: Long-form video understanding requires locating sparse, question-relevant evidence in long, multimodal videos. Real-world video distributions differ in

agentsarxiv-cs-cv
6 Aug 2026
Agents

AgentStream: How Well Do Self-Evolving LLM Agents Perform Under Streaming Tasks?

DGX agent

arXiv:2608.00155v1 Announce Type: cross Abstract: Large language model (LLM) agents can self-evolve by continually improving from their own accumulated experience. However, existing studies predominan

agentsarxiv-cs-lg
5 Aug 2026
Safety

BAP-SQL: Budget-Aware Observation Planning for Agentic Text-to-SQL

DGX agent

arXiv:2608.02876v1 Announce Type: new Abstract: Tool-using agents do not merely consume observations: their actions determine what arrives next. In agentic text-to-SQL, a broad query can spend context

safetyarxiv-cs-ai
5 Aug 2026
Agents

Before Reasoning Can Fail: Pre-Evidence Procedural Failures in Agentic RAG

DGX agent

arXiv:2608.02011v2 Announce Type: replace Abstract: Agentic retrieval-augmented generation (RAG) systems can fail before evidence-conditioned reasoning is tested: an agent may retrieve candidate snipp

agentsarxiv-cs-ai
5 Aug 2026
Agents

TraceCAD: Trace-Guided Repair for Agentic CAD Generation

DGX agent

arXiv:2608.03062v1 Announce Type: new Abstract: LLM-based CAD agents produce executable parametric programs, but their correction loops may lose evidence about satisfied requirements, faulty operation

agentsarxiv-cs-ai
5 Aug 2026
Safety

CRISP: Critical Step Perception for Training Efficient Deep Search Agents

DGX agent

arXiv:2608.01867v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly extended into deep search agents that solve complex questions through multi-step interaction with external

safetyarxiv-cs-cl
4 Aug 2026
Safety

DiffuseAgent-MI: Distributionally-Grounded,Tool-Integrated Self-Evolving Agents for Faithful Visual Reasoning

DGX agent

arXiv:2608.00540v1 Announce Type: new Abstract: Tool-integrated vision-language agents have made remarkable progress on compositional and multi-step visual reasoning. Yet their outputs frequently exhi

safetyarxiv-cs-cv
4 Aug 2026
Agents

MedTextWeaver: Procedural Knowledge Evolution in Agentic Medical Text Editing

DGX agent

arXiv:2602.00740v2 Announce Type: replace Abstract: Medical text editing is essential for improving communication among diverse stakeholders in clinical settings. However, adapting LLM agents to this

agentsarxiv-cs-cl
4 Aug 2026
Agents

When Prompts Control Robots: Prompt Injection Attacks in Multi-Agent Robotic Systems

DGX agent

arXiv:2608.00747v1 Announce Type: new Abstract: Large language models are increasingly integrated into autonomous robotic systems for task planning and control, but this integration exposes them to pr

agentsarxiv-cs-ro
4 Aug 2026
Model Releases

Agentic Harness for Real-World Compilers

DGX agent

arXiv:2603.20075v2 Announce Type: replace-cross Abstract: Compilers are critical to modern computing, yet fixing compiler bugs is difficult. While recent large language model (LLM) advancements enable

model-releasesarxiv-cs-ai
3 Aug 2026
Safety

CAGE: Certified Authorization under Typed-Return Uncertainty for Tool-Using Agents

DGX agent

arXiv:2607.29190v1 Announce Type: new Abstract: Tool-using LLM agents act on typed tool returns, records pairing provenance and categorical fields with numerical values. Runtime permission gates gener

safetyarxiv-cs-ai
3 Aug 2026
Local Ai

Multimodal Reinforcement Learning with Adaptive Verifier for AI Agents

DGX agent

arXiv:2512.03438v3 Announce Type: replace Abstract: Agentic reasoning models trained with multimodal reinforcement learning (MMRL) have become increasingly capable, yet they are almost universally opt

local-aiarxiv-cs-ai
3 Aug 2026
Agents

NeSyFS: A Neuro-symbolic Fast-Slow Thinking Framework for LLM Agent under Partial Observability

DGX agent

arXiv:2607.28942v1 Announce Type: new Abstract: Recently Large Language Models (LLMs) have been increasingly deployed as autonomous agents in applications such as self-reflection, retrieval-augmented

agentsarxiv-cs-ai
3 Aug 2026
Agents

FaithEyes: Towards Faithful Tool Use via Multi-Agent Process-Image Verification

DGX agent

arXiv:2607.28225v1 Announce Type: new Abstract: Agentic vision-language models (VLMs), which interleave textual reasoning with explicit tool calls such as cropping and code-based image manipulation, h

agentsarxiv-cs-cv
31 Jul 2026
Model Releases

ORCA-bench: How Ready Are Language Model Agents for Oncall?

DGX agent

arXiv:2607.28545v1 Announce Type: new Abstract: Large language models can write, patch, and search code, but oncall root cause analysis (RCA) demands something different: reasoning over noisy metrics,

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

SARC-DQ: Runtime Data-Quality Gating for Agentic AI: Silent Evidence Defects, the Incompetence Shield, and Downstream-Only Remediation

DGX agent

arXiv:2607.26313v1 Announce Type: cross Abstract: Agentic systems act, so a defect in the evidence they retrieve becomes a wrong action with a currency cost. The most dangerous enterprise defects are

model-releasesarxiv-cs-ai
31 Jul 2026
Model Releases

TREK: A Travel Reasoning and Evaluation Kit for LLM Agents in Complex Trip Planning

DGX agent

arXiv:2607.26977v1 Announce Type: new Abstract: Travel planning is a demanding stress test for tool-using LLM agents: a usable itinerary is a single artifact that must be right along many axes at once

model-releasesarxiv-cs-cl
30 Jul 2026
Agents

Voice Memory for Agentic Speech Recognition

DGX agent

arXiv:2607.26410v1 Announce Type: new Abstract: We present Voice Memory, a inference-only scheme for agentic speech recognition: at stream time, a frozen corrector reads a single per-domain memory.md

agentsarxiv-cs-cl
30 Jul 2026
Model Releases

Authoring Agent Skills: A Software-Engineering Approach

DGX agent

arXiv:2607.25032v1 Announce Type: cross Abstract: Agent Skills are an emerging way to extend large language model agents with reusable procedural knowledge that the agent loads on demand. Anthropic in

model-releasesarxiv-cs-ai
29 Jul 2026
Agents

A corrective agentic hybrid RAG and an operations-grounded evaluation for a scientific facility

DGX agent

arXiv:2607.24663v1 Announce Type: cross Abstract: Scientific user facilities accumulate decades of operational knowledge that no single search index covers: electronic logbooks, technical documents, i

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

Beyond Sequential Interaction: Benchmarking Parallel Execution and Coordination for GUI Agents

DGX agent

arXiv:2607.22689v1 Announce Type: new Abstract: Graphical user interface (GUI) agents are systems powered by large multimodal models (LMMs). They perceive screen state and execute user instructions th

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Moral Hazard in Multi-Agent Language Models

DGX agent

arXiv:2607.23982v1 Announce Type: cross Abstract: Cooperation can fail when socially valuable effort is costly, weakly observable, and mainly benefits others. Drawing on Holmstrom's team moral-hazard

safetyarxiv-cs-ai
28 Jul 2026
Agents

Multi-agent DRL-based Lane Change Decision Model for Cooperative Platooning in Mixed Traffic

DGX agent

arXiv:2601.11809v2 Announce Type: replace Abstract: Connected automated vehicles (CAVs) possess the ability to communicate and coordinate with one another, enabling cooperative platooning that enhance

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

StateAct: Program State, before Pixels, for Long-Horizon Computer-Use Agents

DGX agent

arXiv:2607.22798v1 Announce Type: cross Abstract: Computer-use agents are usually improved by strengthening perception: better models for reading a screenshot and choosing where to click. Yet a screen

model-releasesarxiv-cs-cv
28 Jul 2026
Agents

Multi-Agent Debate and Visual Information Extraction for SeePhys Pro: A 1st-Place Technical Report from ICML 2026 AI4Math Track 3 Challenge

DGX agent

arXiv:2607.21946v1 Announce Type: new Abstract: This technical report presents our approach to Challenge Track~3: SeePhys Pro at the 3rd AI for Math Workshop, where the task is to answer college-level

agentsarxiv-cs-lg
27 Jul 2026
Model Releases

SwiftMem: Fast Agentic Memory via Query-aware Indexing

DGX agent

arXiv:2601.08160v2 Announce Type: replace Abstract: Agentic memory systems have become critical for enabling LLM agents to maintain long-term context and retrieve relevant information efficiently. How

model-releasesarxiv-cs-cl
27 Jul 2026
Safety

Workload-Aware Caching for Multi-Agent Systems

DGX agent

arXiv:2607.20495v1 Announce Type: new Abstract: Multi-agent systems decompose complex tasks into directed acyclic graphs (DAGs) of specialized agent executions, creating natural opportunities for cach

safetyarxiv-cs-ai
24 Jul 2026
Agents

Dreamer-CPC: Message Learning with World Models for Decentralized Multi-agent Reinforcement Learning

DGX agent

arXiv:2607.19809v1 Announce Type: cross Abstract: In multi-agent reinforcement learning (MARL), inter-agent communication is effective for improving performance under partial observability. Representa

agentsarxiv-cs-lg
23 Jul 2026
Agents

Contract-Grounded Behavior Tree Synthesis via Coding Agents

DGX agent

arXiv:2607.12220v1 Announce Type: new Abstract: Synthesizing deployable robot behavior trees (BTs) from natural language (NL) requires grounding to ensure every generated BT references only skills a r

agentsarxiv-cs-ro
15 Jul 2026
Safety

Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models

DGX agent

arXiv:2607.12463v1 Announce Type: new Abstract: Coding agents must integrate external tool returns into ongoing reasoning - a capability that standard left-to-right pretraining on code exposes only in

safetyarxiv-cs-ai
15 Jul 2026
Model Releases

Lost in the Maze: Overcoming Context Limitations in Long-Horizon Agentic Search

DGX agent

arXiv:2510.18939v2 Announce Type: replace Abstract: Long-horizon agentic search requires iteratively exploring the web over long trajectories and synthesizing information across many sources, enabling

model-releasesarxiv-cs-cl
15 Jul 2026
Agents

Playing ZendoWorld: Challenging AI Agents on Active Visual Concept Induction

DGX agent

arXiv:2607.08233v1 Announce Type: new Abstract: A central challenge in building intelligent systems is enabling agents to jointly perceive complex inputs, form hypotheses about hidden patterns, and de

agentsarxiv-cs-ai
10 Jul 2026
Agents

The Context Access Divide: Interaction-Level Architecture as a Complementary Dimension of Agentic Inequality

DGX agent

arXiv:2607.08495v1 Announce Type: cross Abstract: Sharp et al. (2025) introduce 'agentic inequality' as a framework for analyzing disparities in access to AI agents across three dimensions: availabili

agentsarxiv-cs-ai
10 Jul 2026
Model Releases

UniClawBench: A Universal Benchmark for Proactive Agents on Real-World Tasks

DGX agent

arXiv:2607.08768v1 Announce Type: new Abstract: The rapid development of large language models and multimodal large language models has accelerated the emergence of proactive agents capable of operati

model-releasesarxiv-cs-cl
10 Jul 2026
Agents

Agent-Exploitation Affordances: From Basic to Complex Representation Patterns

DGX agent

arXiv:2607.07475v1 Announce Type: new Abstract: In robotics, the capability of an artificial agent to represent the range of its action possibilities, i.e. affordances, is crucial to understand how it

agentsarxiv-cs-ro
9 Jul 2026
Agents

Physics-Audited Agentic Discovery in Scientific Machine Learning

DGX agent

arXiv:2607.07379v1 Announce Type: new Abstract: In agentic scientific machine learning (SciML), large language model (LLM) agents can discover surrogate models and select one by an automated score, ty

agentsarxiv-cs-ai
9 Jul 2026
Agents

Danus: Orchestrating Mathematical Reasoning Agents with Fact-Graph Memory

DGX agent

arXiv:2607.06447v1 Announce Type: new Abstract: Recent LLM-based mathematical reasoning agents have begun to tackle research-level problems and, in several cases, have contributed to the resolution of

agentsarxiv-cs-ai
8 Jul 2026
Agents

What Do AI Agents Actually Change? An Empirical Taxonomy of Mutation Patterns in Performance-Improving Pull Requests

DGX agent

arXiv:2607.05666v1 Announce Type: cross Abstract: AI coding agents are black boxes: we cannot inspect how they generate code, but we can inspect what they change. This distinction matters for search-b

agentsarxiv-cs-ai
8 Jul 2026
Local Ai

Agentic-V2X: Small Language Model Agents for Deadline-Aware V2X Scheduling in 5G/6G Networks

DGX agent

arXiv:2607.04290v1 Announce Type: cross Abstract: Large Language Models (LLMs) are proposed as control interfaces for next-generation networks, but their latency, hallucinations, and lack of control g

local-aiarxiv-cs-ai
7 Jul 2026
Agents

Compressing the Validation Bottleneck: An Agentic Self-Driving Lab for Scientific Discovery

DGX agent

arXiv:2607.04508v1 Announce Type: new Abstract: Agentic AI-for-Science can automate ideation, planning, and analysis, but final validation still depends on real experiments. A self-driving lab (SDL) c

agentsarxiv-cs-ai
7 Jul 2026
Agents

Self-Supervised Goal-Reaching Results in Multi-Agent Cooperation and Exploration

DGX agent

arXiv:2509.10656v2 Announce Type: replace-cross Abstract: For groups of autonomous agents to achieve a particular goal, they must engage in coordination and long-horizon reasoning. Rather than relying

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

Diverse Evidence, Better Forecasts: Multi-Agent Deliberation Under Information Asymmetry

DGX agent

arXiv:2607.01661v1 Announce Type: new Abstract: Multi-agent systems are increasingly used for forecasting future events, as deliberation among multiple LLMs is believed to improve reasoning and calibr

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Janus: a Playground for User-Involved Agentic Permission Management

DGX agent

arXiv:2607.01510v1 Announce Type: new Abstract: AI agents that autonomously execute tool calls on a user's behalf raise pressing questions about permission management: what role could users play, and

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

Same-Origin Policy for Agentic Browsers

DGX agent

arXiv:2606.14027v3 Announce Type: replace-cross Abstract: Agentic browsers integrate autonomous AI agents into web browsers, enabling users to accomplish web tasks through natural-language instruction

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

CAPTCHA Solving for Native GUI Agents: Automated Reasoning-Action Data Generation and Self-Corrective Training

DGX agent

arXiv:2603.23559v2 Announce Type: replace-cross Abstract: GUI agents are rapidly shifting from multi-module pipelines to end-to-end, native vision-language models (VLMs) that perceive raw screenshots

agentsarxiv-cs-ai
30 Jun 2026
Agents

NaLA: A 3D Native LLM Layout Agent for High-quality 3D Scene Generation

DGX agent

arXiv:2606.29395v1 Announce Type: new Abstract: Recently, Large Language Models (LLMs) have emerged as promising layout agents for 3D scene generation. Existing layout agents still suffer from implaus

agentsarxiv-cs-cv
30 Jun 2026
← Previous
1…2021222324…230
Next →