AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

Benchmarking the Benchmarks: Evaluating Benchmarks for Conversational Agents

DGX agent

arXiv:2608.06329v1 Announce Type: cross Abstract: Task-oriented conversational agents are evaluated using curated or automatically generated benchmarks, yet benchmark quality is rarely assessed. Poor

model-releasesarxiv-cs-ai
7 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

BlockPython: A Process-Aware Agent-Supported Platform for the Transition from Block-Based to Python Programming

DGX agent

arXiv:2608.05716v1 Announce Type: new Abstract: The transition from block-based to text-based programming requires learners to convert visible program structures into abstract textual expressions, whi

agentsarxiv-cs-ai
7 Aug 2026
Model Releases

CASCADE: An Agentic Regulatory Network Framework for Patient-Data-Validated Downstream Perturbation Prediction

DGX agent

arXiv:2608.05359v1 Announce Type: new Abstract: CASCADE is an agentic framework that predicts downstream transcriptional effects of gene perturbation from precomputed ARACNe regulatory networks, expos

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

Causal Episodic Memory for Feedback-Driven Agent Repair

DGX agent

arXiv:2608.05906v1 Announce Type: new Abstract: LLM agents that repair failures often discard successful corrections, forcing later episodes to rediscover similar solutions. We study whether finalized

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

ChainClaw: A Layered Agent Framework for Reliable On-Chain Execution

DGX agent

arXiv:2608.05790v1 Announce Type: new Abstract: General-purpose large language model agents have achieved strong performance on tool-augmented tasks, yet they rely on assumptions break down in blockch

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Contextual Information Policy Optimization for Search Agents

DGX agent

arXiv:2608.06128v1 Announce Type: new Abstract: Search agents extend large language models beyond static parametric memory by enabling them to acquire and use ex ternal evidence during multi-step reas

safetyarxiv-cs-ai
7 Aug 2026
Safety

Multi-Agent Reinforcement Learning for Online Traffic Scheduling in Time-Sensitive Application

DGX agent

arXiv:2608.05346v1 Announce Type: cross Abstract: Time-sensitive networking (TSN) is increasingly integrated into mobile edge computing (MEC) to support applications with stringent latency requirement

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

OrchestraBench: Evaluating Multi-Agent Orchestration Failure Modes, Recovery, and Decomposition Quality

DGX agent

arXiv:2608.05263v1 Announce Type: new Abstract: Multi-agent orchestration frameworks are moving from demos to production, yet benchmarks typically report task accuracy without diagnosing why a pipelin

model-releasesarxiv-cs-ai
7 Aug 2026
Hardware

An Explainable LLM Agent Layer for Open-World Anomaly Detection in Oil Wells

DGX agent

arXiv:2608.04041v1 Announce Type: new Abstract: Open-World Learning (OWL) pipelines for oil well anomaly detection have recently been shown to combine autoencoder-based detection, multiclass classific

hardwarearxiv-cs-lg
6 Aug 2026
Model Releases

Caching for the Future: Scrub Jay Episodic Memory Principles for Agent Memory Systems

DGX agent

arXiv:2608.04746v1 Announce Type: new Abstract: LLM agents that persist across sessions accumulate stored memories whose validity varies enormously by content type, yet existing memory architectures t

model-releasesarxiv-cs-cl
6 Aug 2026
Model Releases

EDATracer: An Agentic Framework for Large-Scale EDA Artifact Analysis

DGX agent

arXiv:2608.04032v1 Announce Type: cross Abstract: Modern chip design relies on electronic design automation (EDA) tools that generate large, heterogeneous artifacts, including source files, scripts, l

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Helping Music Co-Creation Agents 'Listen' Well: Hierarchical Self-Supervised World Models for Understanding and Generation

DGX agent

arXiv:2608.04378v1 Announce Type: cross Abstract: Collaborative music agents need internal representations rich enough to support both understanding and generation, yet flexible enough for a workflow

model-releasesarxiv-cs-lg
6 Aug 2026
Model Releases

Mimir: A Neuro-Symbolic Memory System with Dynamic Grounding for Embodied Agents in Interactive Environments

DGX agent

arXiv:2608.04933v1 Announce Type: new Abstract: Long-horizon embodied task requires agents to act under partial observability while preserving both scene belief and execution progress. Flat histories

model-releasesarxiv-cs-ro
6 Aug 2026
Model Releases

Strategic Evaluation of Planning Strategies for LLM Agents in Cyber-Physical Systems

DGX agent

arXiv:2608.04265v1 Announce Type: cross Abstract: Evaluations of LLM planning agents largely ask whether a task succeeds or a declared plan is followed. In strategic cyber-physical systems, a stronger

model-releasesarxiv-cs-ai
6 Aug 2026
Local Ai

AgenticSCR: An Autonomous Agentic Secure Code Review for Immature Vulnerabilities Detection

DGX agent

arXiv:2601.19138v2 Announce Type: replace-cross Abstract: Secure code review is critical during pre-integration, where Atlassian developers rely on lightweight analysis tools, while deep security asse

local-aiarxiv-cs-ai
5 Aug 2026
Model Releases

Fail-Fast, Restart-Smart: Early Failure Prediction and Restart for SWE Agentic Tasks

DGX agent

arXiv:2608.03222v1 Announce Type: cross Abstract: Software engineering (SWE) agents resolve repository-level issues through long trajectories that grow increasingly expensive as context accumulates. F

model-releasesarxiv-cs-ai
5 Aug 2026
Safety

Learning Clinical-Trial Strategy: Offline Policy Training for Decision Agents

DGX agent

arXiv:2608.03606v1 Announce Type: new Abstract: Clinical development is sequential decision-making under uncertainty, where a sponsor must plan a portfolio of experiments from heterogeneous evidence.

safetyarxiv-cs-ai
5 Aug 2026
Safety

MutMem: Cryptographically Authorized Mutation in Persistent Agent Memory

DGX agent

arXiv:2608.02843v1 Announce Type: cross Abstract: Persistent agent memory must adapt as later outcomes change earlier evidence, yet mutable retrieval weights create an attribution problem: reviewers m

safetyarxiv-cs-ai
5 Aug 2026
Safety

Quo Vadis, World Modeling?

DGX agent

arXiv:2608.02713v1 Announce Type: cross Abstract: Continually improving agents require dynamic interaction feedback beyond static supervision, yet direct real-environment interaction is costly, slow,

safetyarxiv-cs-ai
5 Aug 2026
Agents

RoboReact: Agentic Skill Distillation from Generated Egocentric Videos for Generalizable Whole-Body Manipulation

DGX agent

arXiv:2608.03387v1 Announce Type: new Abstract: Humanoid robots have the potential to perform dexterous manipulation in human environments, yet acquiring diverse and generalizable skills remains costl

agentsarxiv-cs-ro
5 Aug 2026
Model Releases

SeaSlides: Semantic Abstraction Layer for Agentic Slide Generation

DGX agent

arXiv:2608.03298v1 Announce Type: new Abstract: Agentic presentation generation must preserve source content, maintain coherent visual design, render specialized objects, and produce usable artifacts.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

TraceCompiler: Skill-Guided Mining and Compilation of LLM Agent Traces into Mostly Deterministic Workflows

DGX agent

arXiv:2608.02680v1 Announce Type: cross Abstract: Tool-using language-model agents repeatedly rediscover procedures they have already executed, producing traces that mix reusable structure with retrie

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Video-DeepResearch: Towards the Next-Generation Multimodal Deepresearch Agent

DGX agent

arXiv:2608.03979v1 Announce Type: cross Abstract: We introduce Video-DeepResearch (Video-DR), extending multimodal agents from static images to continuous video streams, a setting that demands dense s

model-releasesarxiv-cs-ai
5 Aug 2026
Research

What Language Does and What the Evidence Supports: A Functional Role Taxonomy and Evidence Audit of Language Grounding in Embodied Agents

DGX agent

arXiv:2608.03099v1 Announce Type: new Abstract: Foundation models place language throughout embodied agents, but its presence does not show what it contributes or how well that contribution is grounde

researcharxiv-cs-cl
5 Aug 2026
Model Releases

AOSpec: Action and Observation Co-Speculation for Low-Latency Agent Serving

DGX agent

arXiv:2608.00881v1 Announce Type: new Abstract: Large language model agents increasingly act through stateful tools, yet model generation and environment execution remain serialized at every step. As

model-releasesarxiv-cs-lg
4 Aug 2026
Local Ai

EviSD: Evidence-Conditioned Self-Distillation for Search-Augmented Agents

DGX agent

arXiv:2608.01359v1 Announce Type: new Abstract: Outcome-based reinforcement learning enables search-augmented language agents to learn from verifiable final answers, but its trajectory-level credit ca

local-aiarxiv-cs-cl
4 Aug 2026
Model Releases

Grounding Agentic VLMs with Dedicated Segmentation for Fine-Grained Vehicle Damage Assessment

DGX agent

arXiv:2608.02470v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly deployed as reasoning agents in real-world visual assessment pipelines, yet their spatial grounding remai

model-releasesarxiv-cs-cv
4 Aug 2026
Safety

Look Ahead Before You Distill: Future Trajectory Validation of Teacher Guidance for Agentic On-Policy Distillation

DGX agent

arXiv:2608.01953v1 Announce Type: new Abstract: On-policy distillation (OPD) provides teacher supervision on states visited by the student, reducing the distribution gap between training and inference

safetyarxiv-cs-cl
4 Aug 2026
Safety

MA-HEAD-Net: Adaptive Rule-Guided Multi-Agent DRL for AoI Minimization in UAV-Assisted Emergency Networks

DGX agent

arXiv:2608.01128v1 Announce Type: cross Abstract: In post-disaster scenarios, unmanned aerial vehicles (UAVs) are critical for establishing emergency communication networks. For time-critical rescue m

safetyarxiv-cs-lg
4 Aug 2026
Research

When Replanning Becomes the Bottleneck: Budgeted Replanning for Embodied Agents

DGX agent

arXiv:2608.01428v1 Announce Type: cross Abstract: Embodied agents replan frequently to recover from execution drift, partial observability, and coordination hazards, but each LLM-based replanning call

researcharxiv-cs-lg
4 Aug 2026
Agents

Generative AI in Action: Field Experimental Evidence from Alibaba's Customer Service Operations

DGX agent

arXiv:2603.29888v2 Announce Type: replace-cross Abstract: In collaboration with Alibaba, we study how a generative AI assistant affects service performance in e-commerce after-sales operations. In a l

agentsarxiv-cs-ai
3 Aug 2026
Agents

Know It, Act on It: Investigating Memory Utilization in LLM Personalization

DGX agent

arXiv:2607.29433v1 Announce Type: new Abstract: As large language model (LLM) agents evolve into personalized companions, memory has emerged as a core capability. However, LLMs face a knowledge utiliz

agentsarxiv-cs-cl
3 Aug 2026
Research

Memory Provenance Laundering in LLM Agents: A Non-Amplification Firewall for Persistent Memory

DGX agent

arXiv:2607.29167v1 Announce Type: cross Abstract: Long-term memory lets large language model(LLM) agents reuse prior preferences and work flows, but it also turns untrusted observations into persisten

researcharxiv-cs-ai
3 Aug 2026
Safety

When Unlearning Fails: Reliable Data Deletion under Post-Training in Agent Networks

DGX agent

arXiv:2607.28829v1 Announce Type: cross Abstract: Self-improving federated agent networks keep training after deployment by collecting new trajectories with the current policy and feeding them back in

safetyarxiv-cs-lg
3 Aug 2026
Safety

Functional Cache Grafting: Robust and Rapid Code-Policy Synthesis for Embodied Agents

DGX agent

arXiv:2606.13097v2 Announce Type: replace-cross Abstract: Code-writing large language models (CodeLLMs) generate executable code policies for embodied agents by translating natural language goals and

safetyarxiv-cs-ai
31 Jul 2026
Agents

SkillSmith: Learning to Compose Parametric Skills and Textual Knowledge

DGX agent

arXiv:2607.27497v1 Announce Type: new Abstract: Agentic systems driven by large language models (LLMs) regularly feature two key mechanisms to autonomously solve complex problems: synthesizing text-ba

agentsarxiv-cs-cl
31 Jul 2026
Safety

AgentGFM: A Graph Foundation Model with Node-Agent Information-Flow Control

DGX agent

arXiv:2607.26533v1 Announce Type: new Abstract: Graph Foundation Models (GFMs) aim to learn transferable knowledge from multi-domain graphs and adapt to unseen scenarios. As a fundamental source of re

safetyarxiv-cs-lg
30 Jul 2026
Local Ai

A Control System, a Dataset, and a Recipe for Making Frozen LLM Agents Learn a Domain

DGX agent

arXiv:2607.25415v1 Announce Type: new Abstract: Production LLM agents are increasingly assembled from a frozen model wrapped in a harness: a prompt template, a tool set, a memory/retrieval layer, a pl

local-aiarxiv-cs-ai
29 Jul 2026
Model Releases

Addressable Recall Compaction for Long Context-Window Control in AI Agents

DGX agent

arXiv:2607.25066v1 Announce Type: new Abstract: Long-horizon LLM agents accumulate reasoning traces, actions, and tool observations that can eventually exceed a model's fixed context window. Existing

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Evaluating VLMs for Autonomous Agent-Driven Geometry Clipping Detection in Video Game QA

DGX agent

arXiv:2607.25921v1 Announce Type: cross Abstract: In this work, we study the use of Vision-Language Models (VLMs) for anomaly detection in an agent-driven game Quality Assurance (QA) pipeline focusing

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

GAUGE: Grading Agent-Built Financial Models Without a Golden Answer

DGX agent

arXiv:2607.24889v1 Announce Type: cross Abstract: Financial models combine public disclosures with analyst assumptions to produce forecasts and valuations. While some components can be checked mechani

model-releasesarxiv-cs-ai
29 Jul 2026
Research

MemLens: A Value-Aware Memory Management System with Interactive Analytics for LLM-based Agents

DGX agent

arXiv:2607.25992v1 Announce Type: cross Abstract: Recently, memory management has become a key infrastructure for LLM-based agents, as it directly affects long-horizon reasoning, personalized response

researcharxiv-cs-ai
29 Jul 2026
Safety

Shared Voxel-Map-Based Cooperative Indoor UAV Guidance with a Multi-Agent Soft Actor-Critic Controller

DGX agent

arXiv:2607.25728v1 Announce Type: cross Abstract: This paper presents a cooperative indoor UAV guidance framework that combines a shared voxel-map world model with a multi-agent Soft Actor-Critic (MAS

safetyarxiv-cs-ai
29 Jul 2026
Safety

VetClaw: An Edge-Cloud Multimodal Agentic System for Veterinary Disease Screening

DGX agent

arXiv:2607.26042v1 Announce Type: new Abstract: We present VetClaw, an edge-cloud multimodal agentic system for early veterinary disease screening. VetClaw uses a camera module as an edge sensing devi

safetyarxiv-cs-cv
29 Jul 2026
Agents

Delegation Intelligence in Deep Search: A Controllable Framework for Disentangled Capability Diagnosis

DGX agent

arXiv:2607.23524v1 Announce Type: new Abstract: Deep search is becoming a core capability of modern agent systems, yet it is typically evaluated solely based on end-to-end answer accuracy. This couple

agentsarxiv-cs-ai
28 Jul 2026
Agents

Lexical discovery in unknown environments orchestrated by Large Language Models

DGX agent

arXiv:2607.22591v1 Announce Type: new Abstract: Populations of autonomous agents deployed in unknown environments (e.g. planetary or deep-sea exploration) must develop shared vocabularies to refer to

agentsarxiv-cs-ai
28 Jul 2026
Local Ai

ReflexTrack: A Feedback-Driven Agent for Training-Free Referring Video Object Segmentation

DGX agent

arXiv:2607.24098v1 Announce Type: new Abstract: Referring video object segmentation (RVOS) requires segmenting a target specified by natural language throughout a video. Recent agentic approaches comb

local-aiarxiv-cs-cv
28 Jul 2026
Model Releases

Agentic coding without the cloud: evaluating open-weight large language models on longitudinal data preparation tasks

DGX agent

arXiv:2607.21482v1 Announce Type: new Abstract: Large language models (LLMs) and agents are now widely used tools in code development, with data typically sent to third-party cloud-based models. Their

model-releasesarxiv-cs-ai
24 Jul 2026
← Previous
1…8586878889…236
Next →