AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

TypeGo: An OS Runtime for Embodied Agents

DGX agent

arXiv:2607.05482v1 Announce Type: cross Abstract: Large language models (LLMs) can plan behavior for embodied agents from natural language, but treating the LLM as a request/response oracle on the cri

agentsarxiv-cs-ro
8 Jul 2026
Agents

A Cost-Aware, Paired Protocol for Auditing Dynamic Tool Synthesis in Agentic Video Question Answering

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2607.01469v2 Announce Type: replace Abstract: Agentic Video Question Answering (VideoQA) systems invoke tools during inference, but their tool libraries are fixed, so recurring procedures are re

agentsarxiv-cs-cv
7 Jul 2026
Safety

ARLArena: A Unified Framework for Stable Agentic Reinforcement Learning

DGX agent

arXiv:2602.21534v3 Announce Type: replace Abstract: Agentic reinforcement learning (ARL) has rapidly gained attention as a promising paradigm for training agents to solve complex, multi-step interacti

safetyarxiv-cs-ai
7 Jul 2026
Agents

LLMoxie: Exploring Agentic AI for Scientific Software Development

DGX agent

arXiv:2607.02703v1 Announce Type: cross Abstract: In this paper, we describe LLMoxie, an institutional AI platform whose three-tiered architecture supports multi-cloud and on-premise inference, a Lite

agentsarxiv-cs-ai
7 Jul 2026
Safety

MAD-PINN: A Decentralized Physics-Informed Machine Learning Framework for Safe and Optimal Multi-Agent Control

DGX agent

arXiv:2509.23960v2 Announce Type: replace-cross Abstract: Co-optimizing safety and performance in large-scale multi-agent systems remains a fundamental challenge. Existing approaches based on multi-ag

safetyarxiv-cs-ai
7 Jul 2026
Agents

OptiAgent: End-to-End Optimization Modeling via Multi-Agent Iterative Refinement

DGX agent

arXiv:2607.05346v1 Announce Type: new Abstract: We propose OptiAgent, a multi-agent framework that, given a natural language description of an Operations Research problem, is able to output a solver-r

agentsarxiv-cs-ai
7 Jul 2026
Safety

Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies

DGX agent

arXiv:2607.03423v1 Announce Type: cross Abstract: Modern AI agent implementations such as frontier coding agents chain multiple tools at runtime that create a security surface that per-tool guardrails

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

SpecEyes: Accelerating Agentic Multimodal LLMs via Speculative Perception and Planning

DGX agent

arXiv:2603.23483v2 Announce Type: replace-cross Abstract: Agentic multimodal large language models (MLLMs) (e.g., OpenAI o3 and Gemini Agentic Vision) achieve remarkable reasoning capabilities through

model-releasesarxiv-cs-cl
7 Jul 2026
Agents

Spinning Straw into Gold: Relabeling LLM Agent Trajectories in Hindsight for Successful Demonstrations

DGX agent

arXiv:2607.04235v1 Announce Type: new Abstract: Large language model agents operate in partially observable, long-horizon settings where obtaining supervision remains a major bottleneck. We address th

agentsarxiv-cs-cl
7 Jul 2026
Safety

Strategic Buying Agents

DGX agent

arXiv:2607.04708v1 Announce Type: cross Abstract: Agentic AI is shifting online shopping from search toward delegated purchasing, where autonomous buying agents monitor markets and decide when to buy

safetyarxiv-cs-ai
7 Jul 2026
Agents

TAMA: A Human-AI Collaborative Thematic Analysis Framework Using Multi-Agent LLMs for Clinical Interviews

DGX agent

arXiv:2503.20666v2 Announce Type: replace-cross Abstract: Thematic analysis (TA) is a widely used qualitative approach for uncovering latent meanings in unstructured text data. TA provides valuable in

agentsarxiv-cs-cl
7 Jul 2026
Agents

The 'I Don't Know' Filter: Enhancing Agentic Reliability in Function Calling

DGX agent

arXiv:2607.04034v1 Announce Type: cross Abstract: The language models that underpin agents have seen a rapid rise in performance on function calling benchmarks. However, the metrics used in the traini

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

AgenticRAGTracer: A Hop-Aware Benchmark for Diagnosing Multi-Step Retrieval Reasoning in Agentic RAG

DGX agent

arXiv:2602.19127v2 Announce Type: replace Abstract: With the rapid advancement of agent-based methods in recent years, Agentic RAG has undoubtedly become an important research direction. Multi-hop rea

model-releasesarxiv-cs-cl
3 Jul 2026
Agents

CLAP: Closed-Loop Training, Evaluation, and Release Control for Domain Agent Post-training

DGX agent

arXiv:2607.01846v1 Announce Type: new Abstract: Domain agents often face noisy business data, uncertain post-training gains, offline/application mismatch, and adapter-release risk. This paper presents

agentsarxiv-cs-ai
3 Jul 2026
Agents

ContextNest: Verifiable Context Governance for Autonomous AI Agent

DGX agent

arXiv:2607.02116v1 Announce Type: new Abstract: Autonomous AI agents increasingly depend on external knowledge stores, yet most retrieval pipelines provide relevance without durable guarantees of prov

agentsarxiv-cs-ai
3 Jul 2026
Agents

Leveraging Metamemory Agent for Enhanced Data-Free Code Generation in Large Language Models

DGX agent

arXiv:2501.07892v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown strong performance in automated code generation, with few-shot prompting widely used for its simplicit

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

Safety Testing LLM Agents at Scale: From Risk Discovery to Evidence-Grounded Verification

DGX agent

arXiv:2607.01793v1 Announce Type: new Abstract: LLM agents increasingly perform autonomous actions through external tools, leading to complex and evolving safety risks. However, existing safety testin

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

A Task-State Representation for Long-Horizon Mobile GUI Agents

DGX agent

arXiv:2607.00502v1 Announce Type: new Abstract: While long-horizon mobile GUI agents typically rely on thought-action-observation loops, they struggle to separate persistent task states from transient

agentsarxiv-cs-cl
2 Jul 2026
Agents

Behavior-Adaptive Conversational Agents: Toward a Fluid Personality Framework

DGX agent

arXiv:2607.01034v1 Announce Type: cross Abstract: Large language model (LLM)-based conversational agents (CAs) are now ubiquitous, creating new opportunities for AI-mediated behavior change. Their cap

agentsarxiv-cs-ai
2 Jul 2026
Agents

Cheap Code, Costly Judgment: A Case Study on Governable Agentic Software Engineering

DGX agent

arXiv:2607.01087v1 Announce Type: cross Abstract: Generative AI is shifting software engineering from a practice organized around scarce implementation effort toward one organized around abundant, low

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

From Signals to Structure: How Memory Architecture Drives Language Emergence in LLM Agents

DGX agent

arXiv:2607.00233v1 Announce Type: new Abstract: How do two agents invent a shared language from scratch? In a Lewis signaling game, a sender and receiver must coordinate on a code using only their int

model-releasesarxiv-cs-ai
2 Jul 2026
Agents

SkillSelect-Serve: Budget-Controllable and QoS-Aware Skill Service Recommendation and Composition for Small LLM Agents

DGX agent

arXiv:2607.00011v1 Announce Type: cross Abstract: Reusable skill libraries are becoming important infrastructure for large language model (LLM) agents, yet existing selection methods often treat skill

agentsarxiv-cs-ai
2 Jul 2026
Agents

AI-Assisted Discovery of Convex Relaxations via Dual Agents

DGX agent

arXiv:2606.31182v1 Announce Type: new Abstract: Recent work shows that LLM agents can improve sharp-constant inequalities by searching for extremal constructions, which yield upper bounds. We address

agentsarxiv-cs-ai
1 Jul 2026
Safety

GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation

DGX agent

arXiv:2603.26266v3 Announce Type: replace Abstract: Large vision-language models have endowed GUI agents with strong general capabilities for interface understanding and interaction. However, due to i

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

SimpleSearch-VL: A Simple Recipe for Multimodal Agentic Deep Search

DGX agent

arXiv:2606.31504v1 Announce Type: new Abstract: We present SimpleSearch-VL, an efficient, reliable, and practical framework for multimodal agentic search. Its core idea is to improve the agent's own s

model-releasesarxiv-cs-cv
1 Jul 2026
Agents

ManimAgent: Self-Evolving Multimodal Agents for Visual Education

DGX agent

arXiv:2606.30296v1 Announce Type: new Abstract: Multi-round reflection lets agents built on large language models recover from failures within a single task, but each task remains an isolated episode:

agentsarxiv-cs-ai
30 Jun 2026
Safety

Proof-of-Guardrail in AI Agents and What (Not) to Trust from It

DGX agent

arXiv:2603.05786v2 Announce Type: replace-cross Abstract: As AI agents become widely deployed as online services, users often rely on an agent developer's claim about how safety is enforced, which int

safetyarxiv-cs-ai
30 Jun 2026
Agents

Self-Evolving Agentic Image Restoration via Deliberate Planning and Intuitive Execution

DGX agent

arXiv:2606.28971v1 Announce Type: new Abstract: Real-world image restoration (IR) remains challenging due to complex and coupled degradations. While recent agentic IR frameworks leverage Large Languag

agentsarxiv-cs-cv
30 Jun 2026
Model Releases

Running the Gauntlet: Re-evaluating the Capabilities of Agents Beyond Familiar Environments

DGX agent

arXiv:2606.14397v2 Announce Type: replace Abstract: As agentic systems continue to evolve and are widely deployed in real-world scenarios, there is a growing demand to faithfully evaluate their capabi

model-releasesarxiv-cs-lg
26 Jun 2026
Safety

When Agents Meet Electric Bus Fleet Operations: Pricing Behavior, Trade-offs, and Policy Implications in an Aggregator Framework

DGX agent

arXiv:2606.26400v1 Announce Type: new Abstract: Agentic systems are changing how complex operational tasks are coordinated, introducing a new paradigm for connecting heterogeneous data sources and aut

safetyarxiv-cs-ai
26 Jun 2026
Agents

Membox: Weaving Topic Continuity into Long-Range Memory for LLM Agents

DGX agent

arXiv:2601.03785v3 Announce Type: replace Abstract: Long-term human-agent dialogues are organized by topic continuity: adjacent turns often develop the same goal, plan, problem, or event, while relate

agentsarxiv-cs-cl
25 Jun 2026
Model Releases

Multi-Agent Goal Recognition with Team- and Goal-Conditioned Reinforcement Learning and Factorized Branch-and-Bound

DGX agent

arXiv:2606.25978v1 Announce Type: cross Abstract: Multi-agent goal recognition asks an observer to jointly infer which agents act together and what each team is trying to achieve, so the hypothesis sp

model-releasesarxiv-cs-lg
25 Jun 2026
Agents

Plausible but Wrong: A case study on Agentic Failures in Astrophysical Workflows

DGX agent

arXiv:2604.25345v2 Announce Type: replace Abstract: Agentic AI systems are increasingly being integrated into scientific workflows, yet their behavior under realistic conditions remains insufficiently

agentsarxiv-cs-ai
25 Jun 2026
Safety

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning

DGX agent

arXiv:2606.24428v1 Announce Type: new Abstract: Experience-driven self-evolution is critical for large language model (LLM) agents to improve through open-world interaction. However, existing experien

safetyarxiv-cs-cl
24 Jun 2026
Agents

Grading the Grader: Lessons from Evaluating an Agentic Data Analysis System

DGX agent

arXiv:2606.24839v1 Announce Type: new Abstract: Agentic data analysis systems produce rich outputs, including code, numerical results, and verbal diagnostics. This makes them more challenging to evalu

agentsarxiv-cs-ai
24 Jun 2026
Model Releases

GUI vs. CLI: Execution Bottlenecks in Screen-Only and Skill-Mediated Computer-Use Agents

DGX agent

arXiv:2606.24551v1 Announce Type: new Abstract: Computer-use agents can execute software tasks through either graphical interfaces or programmatic command interfaces, but existing evaluations confound

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

MEMPROBE: Probing Long-Term Agent Memory via Hidden User-State Recovery

DGX agent

arXiv:2606.24595v1 Announce Type: new Abstract: Long-term memory promises LLM agents that grow more capable across sessions, maintaining an accurate, evolving understanding of the user that interactio

model-releasesarxiv-cs-cl
24 Jun 2026
Safety

Policy Gradient with Self-Attention for Model-Free Distributed Nonlinear Multi-Agent Games

DGX agent

arXiv:2509.18371v2 Announce Type: replace-cross Abstract: Multi-agent games in dynamic nonlinear settings are challenging due to the time-varying interactions among the agents and the non-stationarity

safetyarxiv-cs-ro
24 Jun 2026
Model Releases

SAFARI: Scaling Long Horizon Agentic Fault Attribution via Active Investigation

DGX agent

arXiv:2606.24626v1 Announce Type: new Abstract: As autonomous agents tackle increasingly complex multi-step, multi-agent tasks, their execution trajectories have scaled beyond the constraints of even

model-releasesarxiv-cs-ai
24 Jun 2026
Agents

Skills for the future software profession: beyond agentic AI!

DGX agent

arXiv:2606.21894v2 Announce Type: replace-cross Abstract: As coding agents are rapidly changing software engineering, a natural question is: what are the core skills needed by future software engineer

agentsarxiv-cs-ai
24 Jun 2026
Agents

Sol Video Inference Engine: Agent-Native Full-Stack Acceleration Framework for Efficient Video Generation

DGX agent

arXiv:2606.23743v1 Announce Type: cross Abstract: Modern video diffusion models achieve higher generation quality through scaling, but this also increases inference cost. Although many acceleration me

agentsarxiv-cs-ai
24 Jun 2026
Agents

Thinking While Speaking: Inference-Time Knowledge Transfer for Responsive and Intelligent Conversational Voice Agents

DGX agent

arXiv:2511.07397v2 Announce Type: replace Abstract: Voice agents face a fundamental tension: the reasoning, retrieval, and tool use that make foundation models capable are iterative and slow, while co

agentsarxiv-cs-cl
24 Jun 2026
Agents

GRADE: Graph Representation of LLM Agent Dependency and Execution

DGX agent

arXiv:2606.22741v1 Announce Type: new Abstract: Can one graph represent every kind of LLM agent's run? A trace records what each step did, never what it relied on, the state it read, and the results i

agentsarxiv-cs-lg
23 Jun 2026
Agents

PolicyGuard: Towards Test-time and Step-level Adversary (Backdoor) Defense for Reinforcement Learning Agent

DGX agent

arXiv:2606.12896v2 Announce Type: replace Abstract: While real-world applications of reinforcement learning (RL) are becoming increasingly popular, the security of RL systems deserve more attention an

agentsarxiv-cs-lg
23 Jun 2026
Agents

UltraQuant: 4-bit KV Caching for Context-Heavy Agents

DGX agent

arXiv:2606.20474v2 Announce Type: replace Abstract: Context-heavy agents place unusual pressure on the key-value (KV) cache: long prefixes are reused across many short turns, while concurrency determi

agentsarxiv-cs-lg
23 Jun 2026
Safety

APPO: Agentic Procedural Policy Optimization

DGX agent

arXiv:2606.12384v1 Announce Type: cross Abstract: Recent advances in agentic Reinforcement Learning (RL) have substantially improved the multi-turn tool-use capabilities of large language model agents

safetyarxiv-cs-ai
11 Jun 2026
Agents

ISE: An Execution-Grounded Recipe for Multi-Turn OS-Agent Trajectories

DGX agent

arXiv:2606.11520v1 Announce Type: cross Abstract: Training capable OS agents requires data that simultaneously captures structured user intents, multi-turn task delegation, and grounded tool execution

agentsarxiv-cs-ai
11 Jun 2026
Agents

Skill-Augmented AI Agents for Medical Research Analysis: An Exploratory Multi-Model Human Evaluation in an NSCLC Transcriptomic Biomarker Task

DGX agent

arXiv:2606.11830v1 Announce Type: new Abstract: Background. Large language models and AI agents are increasingly used to support biomedical research, but native model outputs may omit key analytical s

agentsarxiv-cs-ai
11 Jun 2026
← Previous
1…3839404142…233
Next →