AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

AutoMOOSE: An Agentic AI for Autonomous Phase-Field Simulation

DGX agent

arXiv:2603.20986v2 Announce Type: replace Abstract: Phase-field modeling links thermodynamics and kinetics to microstructural evolution, but multiphysics frameworks such as MOOSE require expertise to

model-releasesarxiv-cs-ai
10 Aug 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LMM Modality Transfer: A Pre-requisite for Autonomous GIS Agents

DGX agent

arXiv:2608.06948v1 Announce Type: new Abstract: AI models are becoming increasingly adept at understanding and processing spatial information, thereby facilitating agentic problem-solving in spatial t

safetyarxiv-cs-ai
10 Aug 2026
Model Releases

Multi-Agent Forensic Reasoning for Generalizable Deepfake Video Detection

DGX agent

arXiv:2608.06865v1 Announce Type: cross Abstract: The malicious use of generative artificial intelligence to create highly realistic deepfake videos raises serious ethical concerns and poses substanti

model-releasesarxiv-cs-ai
10 Aug 2026
Safety

AppDeltaWorld: Transition-Grounded Delta Code World Model for Mobile GUI Agents

DGX agent

arXiv:2608.05891v1 Announce Type: new Abstract: Mobile GUI agents can operate apps through pixel perception and touch actions, making them a promising interface for collecting and improving long-horiz

safetyarxiv-cs-ai
7 Aug 2026
Local Ai

Communication-Aware Multi-Agent Reinforcement Learning for Decentralized Cooperative UAV Deployment

DGX agent

arXiv:2603.16141v2 Announce Type: replace-cross Abstract: Autonomous Unmanned Aerial Vehicle (UAV) swarms are increasingly used as rapidly deployable aerial relays and sensing platforms, yet practical

local-aiarxiv-cs-lg
7 Aug 2026
Model Releases

FormBharo: Designing and Evaluating a Voice Agent for Conversational Form Filling in Rural India

DGX agent

arXiv:2608.06027v1 Announce Type: cross Abstract: In India, almost every social benefit starts with a form, yet the people who need these benefits most are often unable to read or write. Reaching them

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

From Siloed Algorithms to Compliance-First Agentic Platforms: A Multi-Layered Architecture for Hospital AI Systems

DGX agent

arXiv:2608.06112v1 Announce Type: new Abstract: Hospitals are rapidly adopting artificial intelligence for triage, imaging, scheduling etc., yet most deployments remain isolated point solutions locked

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

Hardware Keystores for AI Agent Signing Workflows: A Zero-Trust MCP Enforcement Architecture

DGX agent

arXiv:2608.06130v1 Announce Type: cross Abstract: AI agents performing cryptographic operations (signing Git commits, authenticating API calls, issuing certificates) currently store private keys in so

model-releasesarxiv-cs-ai
7 Aug 2026
Local Ai

Innovation-Residual Auditing of Autonomous Analysis Agents: Localization, Detection Limits, Error Control, and Identifiability

DGX agent

arXiv:2608.05490v1 Announce Type: new Abstract: Autonomous agents now carry out entire data analyses, selecting cohorts, joining tables, and fitting models with little step-by-step supervision. When s

local-aiarxiv-cs-ai
7 Aug 2026
Model Releases

Robust Native Language Identification through Agentic Decomposition

DGX agent

arXiv:2509.16666v2 Announce Type: replace Abstract: Large language models (LLMs) often achieve high performance in native language identification (NLI) benchmarks by leveraging superficial contextual

model-releasesarxiv-cs-cl
7 Aug 2026
Model Releases

SkillHEX: Improving Agent Skills via Hypothesis-Driven Autonomous Exploration and Exploitation

DGX agent

arXiv:2608.05628v1 Announce Type: new Abstract: Although agent skills equip LLMs with reusable procedural knowledge, manual maintenance suffers from high costs, unscalability, and misalignment. Real-w

model-releasesarxiv-cs-ai
7 Aug 2026
Model Releases

StreamArena: Toward Continuous, Interactive, and Long-Horizon Agentic Streaming Video Understanding

DGX agent

arXiv:2608.05703v1 Announce Type: new Abstract: Deploying autonomous multimodal agents in continuous, real-world environments requires them to ingest unbounded audio-visual streams and maintain hour-s

model-releasesarxiv-cs-cv
7 Aug 2026
Safety

When Privileged Guidance Misaligns: State-Matched Routing and Contextualized Self-Distillation for Multi-Turn Agents

DGX agent

arXiv:2608.05219v1 Announce Type: new Abstract: Privileged on-policy distillation provides dense supervision for multi-turn agents by allowing a synchronized teacher to re-score the student's response

safetyarxiv-cs-ai
7 Aug 2026
Model Releases

When Self-Evolution Backfires: Pre-Commit Gating against Skill Contamination in LLM Agents

DGX agent

arXiv:2608.05810v1 Announce Type: new Abstract: Self-evolving agents accumulate capability by distilling reusable skills from their execution trajectories, but we find this process is not monotonic: p

model-releasesarxiv-cs-ai
7 Aug 2026
Safety

Agentic Reinforcement Learning with Observation-Calibrated Self-Distillation

DGX agent

arXiv:2608.04788v1 Announce Type: cross Abstract: Large language model agents are commonly trained through reinforcement learning with sparse trajectory-level rewards, which offer limited guidance on

safetyarxiv-cs-ai
6 Aug 2026
Model Releases

Argus: A General-Purpose Agentic Runtime for Long-Horizon Reasoning

DGX agent

arXiv:2608.05144v1 Announce Type: new Abstract: Long-horizon reasoning requires an agentic runtime that can persist when evidence supports its current approach and pivot when measurements reveal failu

model-releasesarxiv-cs-ai
6 Aug 2026
Research

EA-Graph: Artifact-Anchored Verification Memory for Coding Agents under Upstream Drift

DGX agent

arXiv:2608.04278v1 Announce Type: cross Abstract: Coding agents increasingly work across sessions, but prose notes can preserve a conclusion without the program state that supported it. After an upstr

researcharxiv-cs-ai
6 Aug 2026
Safety

Guideline-as-Oracle: Zero-Annotation Training of an Ophthalmic Telephone Triage Agent

DGX agent

arXiv:2608.04772v1 Announce Type: cross Abstract: Scaling supervision for multi-turn medical agents is difficult because expert dialogue annotation is costly and clinical conversations are privacy-res

safetyarxiv-cs-ai
6 Aug 2026
Research

Hierarchical Graph Memory for LLM Agents with Path-level Localization and Rewrite

DGX agent

arXiv:2608.05095v1 Announce Type: new Abstract: Agents for long term reasoning require a memory that can be efficiently and effectively updated over time, as new facts and external feedback continue t

researcharxiv-cs-ai
6 Aug 2026
Model Releases

When Does Latent Communication Pay? A Causal Audit of Relayed KV Caches in Multi-Agent LLMs

DGX agent

arXiv:2608.04893v1 Announce Type: cross Abstract: Multi-agent LLM systems relay key--value caches instead of text and credit their gains to exchanged ``latent thoughts''. That credit is a claim about

model-releasesarxiv-cs-ai
6 Aug 2026
Applications

Crayotter: Learning Long-Horizon Video Editing Agents via Group-Relative Preference Backpropagation

DGX agent

arXiv:2608.02694v1 Announce Type: new Abstract: Long-horizon video editing agents receive final-product feedback only after many interdependent decisions. Yet editing quality is subjective, admits mul

applicationsarxiv-cs-cl
5 Aug 2026
Model Releases

DP-MemView: A Memory Interface for Attribute-Level Transcript Privacy in Long-Term LLM Agents

DGX agent

arXiv:2608.03130v1 Announce Type: cross Abstract: Long-term memory enables persistent personalization in LLM agents, but repeated memory-conditioned responses can cumulatively reveal protected attribu

model-releasesarxiv-cs-cl
5 Aug 2026
Model Releases

LeanMem: Simple and Efficient Long-Term Memory for LLM Agents

DGX agent

arXiv:2608.03463v1 Announce Type: new Abstract: Long-term memory is essential for LLM-based agents to sustain interactions and reliably leverage distant history. However, existing memory systems typic

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

MDArena: Evaluating Coding Agents on Realistic Molecular Dynamics Workflows

DGX agent

arXiv:2608.02642v1 Announce Type: cross Abstract: Accelerating scientific discovery is among the most consequential applications of AI, and computational biomolecular simulation stands out as a partic

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

Measurement Without Validity: The Compounding Reliability Problem in Agentic AI Evaluation

DGX agent

arXiv:2608.00794v2 Announce Type: replace Abstract: Agentic AI evaluation pipelines produce benchmark scores that justify deployment decisions, safety certifications, and regulatory compliance claims.

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

MT-Web2Code: Benchmarking Coding Agents on Multi-Turn Regional Reconstruction and Localized Modification

DGX agent

arXiv:2608.03474v1 Announce Type: new Abstract: Recent advances in Large Vision-Language Models (LVLMs) have demonstrated impressive capabilities in web UI generation. However, existing benchmarks pre

model-releasesarxiv-cs-cv
5 Aug 2026
Model Releases

Rethinking Self-Evolving Agent Skills: Feedback Dynamics over Multiple Rounds

DGX agent

arXiv:2608.02636v1 Announce Type: cross Abstract: Self-evolving skill systems promise to improve agents by turning execution feedback into persistent skill updates without changing the underlying mode

model-releasesarxiv-cs-ai
5 Aug 2026
Research

SkillTrace: Traversing a Query-Skill Graph for Composable LLM Agents

DGX agent

arXiv:2608.02356v2 Announce Type: replace Abstract: Large language model agents increasingly solve complex tasks by composing reusable skills from a library. To address this, the key challenge is not

researcharxiv-cs-ai
5 Aug 2026
Model Releases

TARL: Transaction-Aware Reliable Ledgers for Executable Memory Management in Long-Term Agents

DGX agent

arXiv:2608.03699v1 Announce Type: new Abstract: Persistent memory helps long-term agents retain knowledge, yet a single update error can repeatedly distort future retrieval and reasoning. Most existin

model-releasesarxiv-cs-ai
5 Aug 2026
Model Releases

TumorBoard: Evidence-Grounded Multi-Agent Decision Support for Longitudinal Neuro-Oncology

DGX agent

arXiv:2608.03190v1 Announce Type: new Abstract: Neuro-oncology decisions require coordinated interpretation of serial MRI, pathology, molecular markers, treatment history, performance status, and evol

model-releasesarxiv-cs-ai
5 Aug 2026
Local Ai

Verifiable Memory: Learning Unified Memory Management with Local and Global Verifiers for Large Language Model Agents

DGX agent

arXiv:2608.03137v1 Announce Type: new Abstract: Large language model (LLM) agents must retain reusable information, control a bounded active context, and recover earlier evidence during long-horizon i

local-aiarxiv-cs-ai
5 Aug 2026
Safety

ACE-GraphRAG: Agentic Context Engineering for Hierarchical GraphRAG

DGX agent

arXiv:2608.01269v1 Announce Type: new Abstract: Hierarchical Graph Retrieval-Augmented Generation (GraphRAG) organizes corpus knowledge at multiple levels of granularity, yet fixed context constructio

safetyarxiv-cs-cl
4 Aug 2026
Model Releases

Learning Compositional Meta-Routing for Agentic Workflows: An Executable Benchmark

DGX agent

arXiv:2608.00106v1 Announce Type: new Abstract: Agentic systems must decide not only what answer to produce, but which reasoning and execution operations should precede it. A controller may answer dir

model-releasesarxiv-cs-lg
4 Aug 2026
Model Releases

MetaRoute-Bench: Evaluating Meta-Decision Policies for Agentic Workflow Routing

DGX agent

arXiv:2608.00107v1 Announce Type: new Abstract: Agentic systems must repeatedly decide whether to answer directly, decompose a task, invoke a tool, execute code, delegate to a specialist, verify an in

model-releasesarxiv-cs-lg
4 Aug 2026
Research

Beyond Retrieval: Analytic Memory for Multimodal Agents

DGX agent

arXiv:2607.29440v1 Announce Type: new Abstract: Long-term multimodal memory must support not only retrieving relevant information but also computing over observations accumulated across interactions.

researcharxiv-cs-ai
3 Aug 2026
Safety

Fidelity Is Not Safety: Gently-Compressed LLMs Pass Every Data-Free Quality Guard Yet Invent Procedure Steps in Agentic Execution

DGX agent

arXiv:2607.28196v1 Announce Type: new Abstract: Practitioners accept a compressed language model once it clears a stack of data-cheap quality guards: perplexity within a small factor of the original,

safetyarxiv-cs-cl
31 Jul 2026
Model Releases

Flat Score, Amplified Failures: How the Error Budget Masks Damage in Quantized LLM Agents

DGX agent

arXiv:2607.27275v1 Announce Type: new Abstract: Post-training quantization to 4-bit weights is widely reported to be nearly lossless. We test this claim for multi-turn, tool-calling agents, where it n

model-releasesarxiv-cs-lg
31 Jul 2026
Safety

Graph Is the Verifier: Agentic Reinforcement Learning for Interprocedural Vulnerability Detection

DGX agent

arXiv:2607.26656v1 Announce Type: cross Abstract: Real-world vulnerabilities often span multiple functions, yet most learning-based detectors classify each function in isolation: on a sample of real C

safetyarxiv-cs-ai
31 Jul 2026
Model Releases

GuideSkill: Evolving Executable LLM Agent Skills for Guideline-Grounded Clinical Reasoning

DGX agent

arXiv:2607.26160v1 Announce Type: new Abstract: Clinical practice guidelines (CPGs) encode diagnostic criteria, but LLM systems typically retrieve guideline text or absorb it through training rather t

model-releasesarxiv-cs-ai
31 Jul 2026
Safety

HALO: Heterogeneous Admission through Localized Obligations for Safe Agentic Execution

DGX agent

arXiv:2607.27636v1 Announce Type: cross Abstract: Recent agentic AI systems may return a heterogeneous response containing notices, requests, handoffs, and actions. Conditions can change before extern

safetyarxiv-cs-ro
31 Jul 2026
Model Releases

MagicSelector: Joint Optimization for Agent Tool Selection via Counterfactual Decomposition and Progressive Reranking

DGX agent

arXiv:2607.17751v2 Announce Type: cross Abstract: We present MagicSelector, a joint optimization framework integrating Counterfactual task decomposition, Progressive reranking, and Dynamic Top-K, desi

model-releasesarxiv-cs-cl
31 Jul 2026
Safety

CinemaTraj: Composing Atomic Camera Trajectories for 3D Scenes with LLM Agents

DGX agent

arXiv:2607.26910v1 Announce Type: new Abstract: Automatically generating cinematically expressive camera trajectories through 3D scenes from natural language descriptions is a challenging task of high

safetyarxiv-cs-cv
30 Jul 2026
Model Releases

EvoPINN: Agentic Discovery of Executable Algorithms for Physics-Informed Neural Networks

DGX agent

arXiv:2607.26490v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) have emerged as a powerful paradigm for solving partial differential equations (PDEs), yet their performance

model-releasesarxiv-cs-lg
30 Jul 2026
Local Ai

Think Short, Defer Smart, Act, and Repeat: Calibrated Reasoning and Uncertainty-Aware Deferral for Edge LLM Agents

DGX agent

arXiv:2607.26865v1 Announce Type: cross Abstract: LLM agents following the ReAct paradigm are promising enablers of complex multi-step tasks, including multi-hop question answering, code generation, a

local-aiarxiv-cs-lg
30 Jul 2026
Safety

Explanation-Bound Tool Execution for AI Agents: Server-Verified Action Claims Without Trusting Model Rationales

DGX agent

arXiv:2607.25364v1 Announce Type: new Abstract: Tool-using agents expose structured calls but commonly attach free-form rationales. Such rationales are neither authorization nor reliable introspection

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

HOBA: Hierarchical On-Policy Bidding Agents for Adaptive Online Advertising

DGX agent

arXiv:2607.24779v1 Announce Type: new Abstract: Online advertising bidding systems typically deploy multiple offline-trained expert models (e.g., PID controllers, model predictive control, offline RL

model-releasesarxiv-cs-ai
29 Jul 2026
Safety

Interpretable GOHR Agents via Sparse Autoencoders

DGX agent

arXiv:2607.25132v1 Announce Type: new Abstract: A central challenge in interpreting learned decision-making systems is to determine whether their internal representations contain concepts that help ex

safetyarxiv-cs-lg
29 Jul 2026
Model Releases

Kernel Forge: An Agent Harness for LLM-based Generation and Optimization of CUDA Kernels

DGX agent

arXiv:2607.24762v1 Announce Type: new Abstract: Machine learning models are increasingly embedded in everyday software, and most of their runtime is spent in a small set of compute kernels such as mat

model-releasesarxiv-cs-ai
29 Jul 2026
← Previous
1…96979899100…236
Next →