AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

Efficient Reinforcement Learning for Long-Horizon Tool-Use Agentic Tasks

DGX agent

arXiv:2608.10357v1 Announce Type: cross Abstract: Long-horizon tool-using agents must reason over user goals, domain policies, tool calls, simulator state, and delayed verifiable rewards. Reinforcemen

model-releasesarxiv-cs-ai
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Agentic Router: An Execution-Grounded Continual Learning Approach With Memory

DGX agent

arXiv:2608.09184v1 Announce Type: new Abstract: Large language model (LLM) agents provide a promising interface for command-line-based network operations, but a plausible command may still fail or int

agentsarxiv-cs-ai
11 Aug 2026
Safety

Artificial Leviathan: Exploring Social Evolution of LLM Agents Through the Lens of Hobbesian Social Contract Theory

DGX agent

arXiv:2406.14373v3 Announce Type: replace Abstract: The emergence of Large Language Models (LLMs) and advancements in Artificial Intelligence (AI) offer an opportunity for computational social science

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Beyond the Capability Boundary: Zeroth-Order Optimization for Self-Evolving LLM Agents

DGX agent

arXiv:2608.09292v1 Announce Type: cross Abstract: Self-evolving methods improve the capabilities of LLM agents by sampling trajectories from the underlying LLMs and learning from these trajectories. H

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

Business Truth, not SQL Accuracy: A Rule-Gated 7B Analytics Agent Outperforms a Direct-Prompted 32B Baseline

DGX agent

arXiv:2608.09254v1 Announce Type: new Abstract: LLM analytics agents are evaluated on SQL syntax accuracy, but production failures look different: questions with two valid business definitions, questi

agentsarxiv-cs-ai
11 Aug 2026
Agents

Compositional Threat Analysis of Latent Compromise in LLM Agent Systems: The Order 66 Scenario

DGX agent

arXiv:2608.08131v1 Announce Type: cross Abstract: In the fictional Order 66, catastrophe does not arise from a powerful command alone: a trusted population is preconditioned, a short directive activat

agentsarxiv-cs-ai
11 Aug 2026
Agents

LAVE: Latent Visual Evidence-Enhanced Planning for Video Tool-use Agents

DGX agent

arXiv:2608.07585v1 Announce Type: new Abstract: Long-video understanding requires models to efficiently acquire and reuse sparse visual evidence from long and redundant video streams. Recent video too

agentsarxiv-cs-cv
11 Aug 2026
Agents

NeuroRefiner: Morphology-Aware Multi-Agent Refinement for 3D Fluorescence Microscopy Neuron Segmentation

DGX agent

arXiv:2608.09636v1 Announce Type: cross Abstract: Accurate 3D neuron segmentation in fluorescence microscopy is critical for neuroscience. However, the sparse and elongated morphology of neurons poses

agentsarxiv-cs-ai
11 Aug 2026
Agents

Preference Redirection via Attention Concentration: An Attack on Computer Use Agents

DGX agent

arXiv:2604.08005v2 Announce Type: replace Abstract: Advancements in multimodal foundation models have enabled the development of Computer Use Agents (CUAs) capable of autonomously interacting with GUI

agentsarxiv-cs-lg
11 Aug 2026
Model Releases

SHE: Trajectory-driven Safety Harness Evolution for LLM Agents

DGX agent

arXiv:2608.09885v1 Announce Type: new Abstract: The safety of large language model (LLM) agents depends not only on model weights but also on the agent harness that manages context, memory, tools, per

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

An Agentic Hybrid Top-Down and Bottom-Up Approach to Knowledge Graph Generation

DGX agent

arXiv:2608.07023v1 Announce Type: cross Abstract: Organizing thousands of unstandardized, multilingual expertise declarations is a persistent challenge for Human Resources (HR) platforms, directly imp

agentsarxiv-cs-ai
10 Aug 2026
Safety

DiDPO: Diff-in-Diff Policy Optimization for Coding Agent Training

DGX agent

arXiv:2608.07147v1 Announce Type: new Abstract: Reinforcement learning with Verifiable Reward (RLVR) has emerged as a powerful paradigm for training coding agents, where the execution feedback from co

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

HarnessSafe: Evaluating Safety Across Persistent Carriers in Agent Harnesses

DGX agent

arXiv:2608.06984v1 Announce Type: cross Abstract: Modern agent harnesses persist state across tasks and sessions through persistent carriers like memory, skills, tools, and shared artifacts. However,

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

MemPrism: Task-Conditioned Relational Memory Views for Long-Horizon Agents

DGX agent

arXiv:2608.06745v1 Announce Type: new Abstract: Long-horizon agents rely on memory to reuse experiences, yet existing memory systems often assume that evidence can be directly consumed through a fixed

safetyarxiv-cs-ai
10 Aug 2026
Agents

Risk-Aware Decision Policies for Agents Under Noisy Perception

DGX agent

arXiv:2608.06420v1 Announce Type: cross Abstract: Perception in biological systems is inherently noisy, requiring organisms to make decisions under uncertainty where misclassification can be costly or

agentsarxiv-cs-ai
10 Aug 2026
Research

TEPA: Revoking Stale Memories for Conflict-Robust Language Agents

DGX agent

arXiv:2608.07429v1 Announce Type: new Abstract: Long-term memory enables language agents to reuse past facts, preferences, and task experience. Persistence also creates a central falsifiability proble

researcharxiv-cs-ai
10 Aug 2026
Agents

Toward Reliable Context Compression for Long-Horizon Agents: An Empirical Study of Execution Instability

DGX agent

arXiv:2608.06503v1 Announce Type: new Abstract: Recurrent context compression controls context growth in long-horizon agents, but its behavioral effects remain poorly understood. In this preliminary e

agentsarxiv-cs-lg
10 Aug 2026
Agents

Beyond Top-K: Replacing Black-Box Retrieval with Interpretable Agentic Operations

DGX agent

arXiv:2608.06305v1 Announce Type: new Abstract: Retrieval-augmented generation over long documents is dominated by one design: chunk the text, embed the chunks, and surface the top-k nearest neighbour

agentsarxiv-cs-ai
7 Aug 2026
Safety

EvoHarness-RL: Learning Self-Evolving Runtime Harness for Long-Horizon LLM Agents

DGX agent

arXiv:2608.05446v1 Announce Type: cross Abstract: Long-horizon LLM agents increasingly rely on external execution support to maintain state, track progress, invoke tools, verify outcomes, and reuse ex

safetyarxiv-cs-cl
7 Aug 2026
Agents

F^2Agent: Financial Fusion of Agentic Intelligence for Multimodal Trading

DGX agent

arXiv:2608.05668v1 Announce Type: cross Abstract: With increasingly diverse and heterogeneous information sources, effectively leveraging multimodal data is becoming pivotal for high-quality financial

agentsarxiv-cs-ai
7 Aug 2026
Agents

Multi-Agent Transformer for Queue-Level XR Traffic Scheduling in TSN Networks

DGX agent

arXiv:2608.05340v1 Announce Type: cross Abstract: Time-Sensitive Networking (TSN) and Mobile Edge Computing (MEC) hold strong potential for enabling ultra-reliable low-latency communication for time-s

agentsarxiv-cs-ai
7 Aug 2026
Agents

OPERA: Operator-residual feedback for reliable autonomous optical experiments with language-model agents

DGX agent

arXiv:2608.05990v1 Announce Type: new Abstract: Autonomous agents choose actions using scores that may not reflect experimental success. We developed OPERA, an operator-residual framework for optical

agentsarxiv-cs-ai
7 Aug 2026
Agents

QuanTiMedAI: Quantum-Enhanced Time-Series Model guided by Agentic AI for Cardiac Arrest Mortality Prediction

DGX agent

arXiv:2608.06294v1 Announce Type: new Abstract: Cardiac arrest remains one of the most lethal conditions encountered in intensive care units. Despite the growing availability of electronic health reco

agentsarxiv-cs-ai
7 Aug 2026
Safety

Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents

DGX agent

arXiv:2608.06353v1 Announce Type: cross Abstract: We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle th

safetyarxiv-cs-ai
7 Aug 2026
Agents

SkillTrace: Multi-Trace Provenance Auditing for LLM-Agent Skill Reuse

DGX agent

arXiv:2608.05204v1 Announce Type: new Abstract: LLM-agent ecosystems are rapidly growing around reusable skills: mixed-modality packages of metadata, natural-language instructions, code, tools, refere

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

SkillTV-Bench: Benchmarking How Well Judges Perform on Skill-Augmented Agentic Execution

DGX agent

arXiv:2608.05573v1 Announce Type: new Abstract: LLM agents increasingly execute long-horizon tasks through tool use and environment interaction, shifting evaluation from final-response scoring to veri

model-releasesarxiv-cs-ai
7 Aug 2026
Agents

SkillZip: Contract-Preserving Graph Compression for Scalable Agent Skill Libraries

DGX agent

arXiv:2608.05604v1 Announce Type: cross Abstract: Large Language Models (LLMs) increasingly act as agents whose procedural knowledge is stored in reusable skill packages and loaded at inference time.

agentsarxiv-cs-ai
7 Aug 2026
Agents

When Experience Becomes Instruction: Trajectory Poisoning in Self-Evolving Agent Skill Systems

DGX agent

arXiv:2608.05563v1 Announce Type: cross Abstract: Self-evolving skill (SES) systems distill agent trajectories into persistent skills, allowing untrusted experience to become trusted instruction. We i

agentsarxiv-cs-ai
7 Aug 2026
Agents

EmpaAva: An Open-source Agentic 3D-Avatar Empathetic Live Chatbot

DGX agent

arXiv:2608.04709v1 Announce Type: new Abstract: This paper presents EmpaAva, to our knowledge the first open-source, agentic 3D-avatar empathetic chatbot, which carries empathetic response generation

agentsarxiv-cs-cl
6 Aug 2026
Local Ai

EvolveNet: Collaborative Harness Evolution for Agent Self-Improvement

DGX agent

arXiv:2608.04968v1 Announce Type: new Abstract: The capabilities of an LLM agent depend not only on its model but on the harness: the executable program that constructs context, invokes tools, verifie

local-aiarxiv-cs-lg
6 Aug 2026
Model Releases

FinPerMA: A Theory-Informed, Event-Grounded Personalized-Memory Benchmark for LLM Agents

DGX agent

arXiv:2608.04095v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly used as personalized assistants in high-stakes domains such as financial advising, yet it remains unc

model-releasesarxiv-cs-ai
6 Aug 2026
Safety

Governing Execution Risk in Agentic AI Systems: A Trajectory-Guided Framework for Red Teaming

DGX agent

arXiv:2608.04018v1 Announce Type: cross Abstract: AI agents are increasingly embedded in organizational workflows, where they interact with external information sources and invoke digital tools to per

safetyarxiv-cs-ai
6 Aug 2026
Agents

TopoChunker: Topology-Aware Agentic Document Chunking Framework

DGX agent

arXiv:2603.18409v2 Announce Type: replace Abstract: Current document chunking methods for Retrieval-Augmented Generation (RAG) typically linearize text. This forced linearization strips away intrinsic

agentsarxiv-cs-cl
6 Aug 2026
Agents

TourSynbio-Search: A Large Language Model Driven Agent Framework for Unified Search Method for Protein Engineering

DGX agent

arXiv:2411.06024v1 Announce Type: cross Abstract: The exponential growth in protein-related databases and scientific literature, combined with increasing demands for efficient biological information r

agentsarxiv-cs-ai
6 Aug 2026
Safety

Agentic Reinforcement Learning with Self-Distilled Reward Shaping

DGX agent

arXiv:2608.03223v1 Announce Type: cross Abstract: Agentic reinforcement learning enables LLM agents to learn through interaction, but sparse trajectory-level rewards reveal success without identifying

safetyarxiv-cs-ai
5 Aug 2026
Agents

Beyond Simply Environment Scaling: Designing Effective Environment Distributions for Multimodal Agent Learning

DGX agent

arXiv:2608.03571v1 Announce Type: new Abstract: Recent works train agents by constructing large-scale multimodal environment pools. However, we find that simply increasing the number of multimodal env

agentsarxiv-cs-cv
5 Aug 2026
Model Releases

DataSpace: Benchmarking Data Agents for Verifiable Analytics over Heterogeneous Workspaces

DGX agent

arXiv:2608.03451v1 Announce Type: new Abstract: Data agents enable natural-language analytics over organizational workspaces, where relevant evidence may be scattered across databases, structured file

model-releasesarxiv-cs-ai
5 Aug 2026
Agents

MAFIA: Query-Only Memory Attacks via Probing and Factual Injection against Audited LLM Agents

DGX agent

arXiv:2608.03844v1 Announce Type: new Abstract: Memory-augmented LLM agents rely on rich context for long-horizon reasoning and acting, yet their memory modules expose a persistent attack surface for

agentsarxiv-cs-ai
5 Aug 2026
Safety

Relational Priors as Convergence Pressure in LLM-Based Multi-Agent Systems

DGX agent

arXiv:2608.03239v1 Announce Type: new Abstract: Large language model-based multi-agent systems (LLM-MAS) are designed through roles, debate protocols, and aggregation rules. These choices create impli

safetyarxiv-cs-cl
5 Aug 2026
Safety

A Forward-Inverse Dynamic Game Framework for Enhanced Multi-Agent Trajectory Planning

DGX agent

arXiv:2608.01636v1 Announce Type: new Abstract: This paper studies feedback Nash equilibrium (FBNE) seeking for multi-agent trajectory planning in nonlinear dynamical systems with unknown agents' obje

safetyarxiv-cs-ro
4 Aug 2026
Safety

Quick on the Uptake: Eliciting Implicit Intents from Human Demonstrations for Personalized Mobile-Use Agents

DGX agent

arXiv:2508.08645v3 Announce Type: replace Abstract: As multimodal large language models advance rapidly, the automation of mobile tasks has become increasingly feasible through the use of mobile-use a

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

SERL-SQL: Selective Hindsight Distillation for Text-to-SQL Reinforcement Agentic Learning

DGX agent

arXiv:2608.00485v1 Announce Type: new Abstract: Recent Text-to-SQL systems increasingly rely on multi-turn interaction, execution feedback, and reinforcement learning. However, most existing methods u

model-releasesarxiv-cs-cl
4 Aug 2026
Agents

When Collaboration Becomes a Trigger: Collective Evidence-Threshold Backdoors in Multi-Agent Systems

DGX agent

arXiv:2608.01085v1 Announce Type: cross Abstract: LLM-based multi-agent systems (MAS) extend LLM capabilities through iterative communication and shared contexts. However, this collaboration introduce

agentsarxiv-cs-lg
4 Aug 2026
Agents

AREA3D: Active Reconstruction Agent with Unified Feed-Forward 3D Perception and Vision-Language Guidance

DGX agent

arXiv:2512.05131v2 Announce Type: replace-cross Abstract: Active 3D reconstruction enables an agent to autonomously select viewpoints to efficiently obtain accurate and complete scene geometry, rather

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

Multi-Agent Planning with Spatio-Temporal and Topological Constraints using STL-GO

DGX agent

arXiv:2607.28679v1 Announce Type: new Abstract: Multi-agent planning problems arise in a variety of engineering applications, such as multi-robot wildfire fighting and unmanned aerial inspection in fa

model-releasesarxiv-cs-ai
3 Aug 2026
Model Releases

Beyond Similarity: Grounded Agentic Extraction and Expert-Adjudicated Evaluation of Intertextuality in Classical Chinese Histories

DGX agent

arXiv:2607.27595v1 Announce Type: new Abstract: Computational approaches to intertextuality have advanced from string matching to neural retrieval, yet their outputs, similarity scores and parallel-pa

model-releasesarxiv-cs-cl
31 Jul 2026
Model Releases

Can Agents Deceive? Evaluating Reasoning and Deception in ParliamentBench using a Social Deduction Game

DGX agent

arXiv:2607.28146v1 Announce Type: new Abstract: As large language models (LLMs) are deployed as agents in high-stakes settings, such as medical and legal systems, understanding their deceptive capabil

model-releasesarxiv-cs-cl
31 Jul 2026
Agents

(EC)2: Event-Centric Explainability for Cybersecurity Through Multi-Agent LLM Investigations

DGX agent

arXiv:2607.26201v1 Announce Type: cross Abstract: Security operations centers rely on anomaly detection systems to flag suspicious events. Feature-level explanations for anomaly detectors offer limite

agentsarxiv-cs-ai
31 Jul 2026
← Previous
1…4445464748…233
Next →