AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

SAGE: Stochastic Prompt Optimization via Agent-Guided Exploration

DGX agent

arXiv:2606.18902v2 Announce Type: replace Abstract: Context engineering has emerged as a primary lever for improving AI systems without parameter updates. Recent work showing that textual gradients do

model-releasesarxiv-cs-cl
29 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

AutoMat: Enabling Automated Crystal Structure Reconstruction from Microscopy via Agentic Tool Use

DGX agent

arXiv:2505.12650v2 Announce Type: replace-cross Abstract: Reconstructing atomistic crystal structures from a single noisy STEM projection is an ill-posed inverse problem: multiple lattices can explain

model-releasesarxiv-cs-ai
28 Jul 2026
Hardware

Energy Constrained Hierarchical Underwater Monitoring via Local Multi-Agent RAG

DGX agent

arXiv:2607.24313v1 Announce Type: cross Abstract: Marine life monitoring is limited by strict energy constraints, poor underwater connectivity, and the high cost of transmitting raw multimodal data fr

hardwarearxiv-cs-cv
28 Jul 2026
Local Ai

VecTree-RAG: An Agentic Retrieval-Augmented Generation Framework Combining Vector and Tree Retrieval for Efficiency and Accuracy

DGX agent

arXiv:2607.23006v1 Announce Type: cross Abstract: Scientific question answering requires a retrieval system to solve two distinct problems: identifying which papers are relevant and locating the suppo

local-aiarxiv-cs-ai
28 Jul 2026
Safety

AgentHOI: Multi-Agent Reasoning for Human-Object-Interaction Video Generation via Implicit Representation Alignment

DGX agent

arXiv:2607.22241v1 Announce Type: new Abstract: Recent advances in video diffusion models have spurred interest in human-object interaction (HOI) video generation, which demands fine-grained control o

safetyarxiv-cs-cv
27 Jul 2026
Local Ai

Auditing Provenance Sensitivity in LLM Agent Action Selection

DGX agent

arXiv:2607.20827v1 Announce Type: new Abstract: LLM agents choose tools and arguments from context that mixes user requests, tool outputs, retrieved records, memory, and untrusted text. Evidence can b

local-aiarxiv-cs-ai
24 Jul 2026
Safety

Compile, Then Page: Executable SOP Programs and a Capability-Gated Runtime for Procedural LLM Agents

DGX agent

arXiv:2607.11346v3 Announce Type: replace Abstract: Enterprise agents must follow long-horizon, conditional, safety-critical standard operating procedures (SOPs). We compile machine-readable SOP const

safetyarxiv-cs-ai
24 Jul 2026
Safety

PATS: Policy-Aware Training Scaffolding for Agentic Reinforcement Learning

DGX agent

arXiv:2607.21419v1 Announce Type: new Abstract: In long-horizon LLM agent reinforcement learning, weak policies often repeat similar failures, producing uninformative rollout trajectories and limiting

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

PromptPack: Scaling LLM Annotation Agents for Online Recommendation

DGX agent

arXiv:2607.20528v1 Announce Type: new Abstract: Online recommendation platforms increasingly use Large Language Models (LLMs) to extract structured features from ad creatives. While deploying a single

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Same Dangerous Objective, Opposite Advice: Direct Exposure versus Multi-Agent Mediation

DGX agent

arXiv:2607.21518v1 Announce Type: new Abstract: Even a current high-capability LLM can appear safer when shown a dangerous objective directly than when other agents transform and relay its direction.

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

SciExplore: Evaluating Autonomous Agents from Scientific Navigation to Information Integration

DGX agent

arXiv:2607.20926v1 Announce Type: new Abstract: Scientific research involves complex information-seeking and reasoning workflows across heterogeneous sources. However, existing benchmarks primarily em

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

Self-Evolving Recommendation System: End-To-End Autonomous Model Optimization With LLM Agents

DGX agent

arXiv:2602.10226v2 Announce Type: replace-cross Abstract: Optimizing large-scale machine learning systems, such as recommendation models for global video platforms, requires navigating a massive hyper

model-releasesarxiv-cs-ai
24 Jul 2026
Local Ai

In-the-Flow Agentic System Optimization for Effective Planning and Tool Use

DGX agent

arXiv:2510.05592v2 Announce Type: replace Abstract: Outcome-driven reinforcement learning has advanced reasoning in large language models (LLMs), but prevailing tool-augmented approaches train a singl

local-aiarxiv-cs-ai
23 Jul 2026
Model Releases

Profile-Graph Memory for LLM Agents: Implicit Cross-Entity Traversal through Narrative Profiles

DGX agent

arXiv:2607.19359v1 Announce Type: new Abstract: Long-term memory is essential for LLM agents that interact across sessions, yet current memory benchmarks primarily evaluate single-hop recall, leaving

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

SPINE: Bridging the Cyber-Physical Gap with Agentic AI

DGX agent

arXiv:2607.13049v1 Announce Type: new Abstract: Foundation models have given robots a sophisticated brain for complex decision-making, yet deploying that intelligence into a physical platform still de

model-releasesarxiv-cs-ai
16 Jul 2026
Safety

Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering

DGX agent

arXiv:2603.28583v2 Announce Type: replace-cross Abstract: Despite the success of Vision-Language Models (VLMs), misleading charts remain a significant challenge due to their deceptive visual structure

safetyarxiv-cs-ai
15 Jul 2026
Model Releases

SeqGPT: A Constrained Transformer Agent for the Inverse Designof Multi-Panel Composite Structures

DGX agent

arXiv:2607.11910v1 Announce Type: cross Abstract: Optimizing composite stacking sequences to match continuous targets (e.g., Lamination or Buckling Parameters) with discrete manufacturing constraints

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

A Reliability Assessment of LALM Audio Judges for Full-Duplex Voice Agents

DGX agent

arXiv:2607.07985v1 Announce Type: cross Abstract: We report the empirical reliability of Gemini models as audio judges that score full-duplex agent conversations directly from the raw stereo waveform,

model-releasesarxiv-cs-ai
10 Jul 2026
Safety

Feedback Manipulation Regularization: Enabling Offline Agent Alignment for Imitation Learning

DGX agent

arXiv:2607.07859v1 Announce Type: new Abstract: Reinforcement learning (RL) research has increasingly shifted focus towards alignment, ensuring agents learn behaviors adhering to human values. While h

safetyarxiv-cs-ai
10 Jul 2026
Model Releases

Flow-ERD: Agent-type Aware Flow Matching with Entropy-Regularized Distillation for Diverse Traffic Simulation

DGX agent

arXiv:2607.06957v1 Announce Type: cross Abstract: Realistic and diverse traffic simulation is essential to autonomous driving development. Yet prevailing benchmarks predominantly reward realism, and r

model-releasesarxiv-cs-lg
9 Jul 2026
Model Releases

Multi-Agent Robotic Control with Onboard Vision-Language Models

DGX agent

arXiv:2607.07403v1 Announce Type: cross Abstract: Vision Language Models (VLMs) and Vision Language Action (VLA) models have shown promise in robotic control. Yet, they face significant challenges reg

model-releasesarxiv-cs-ro
9 Jul 2026
Safety

The Blind Curator: How a Biased Judge Silently Disables Skill Retirement in Self-Evolving Agents

DGX agent

arXiv:2607.07436v1 Announce Type: new Abstract: A self-evolving agent retires its bad skills by watching them fail, so what happens when the judge cannot see the failures? Skill retirement is the stru

safetyarxiv-cs-ai
9 Jul 2026
Model Releases

FirstResearch: Auditable Question Formation for LLM Scientific Discovery Agents

DGX agent

arXiv:2607.05682v1 Announce Type: new Abstract: LLM systems for scientific discovery increasingly assist with ideation, literature synthesis, experiment planning, and report generation, but the first

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Harnessing Code Agents for Automatic Software Verification

DGX agent

arXiv:2607.06341v1 Announce Type: cross Abstract: Formal verification offers the strongest guarantee of software correctness, but it does not scale: the proofs demanded by interactive theorem provers

model-releasesarxiv-cs-ai
8 Jul 2026
Model Releases

Onnes: A Physics-Grounded Multi-Agent LLM Simulator for Cryogenic Fault Diagnosis in Quantum Computing Infrastructure

DGX agent

arXiv:2607.05805v1 Announce Type: new Abstract: Dilution refrigerators are the enabling infrastructure of superconducting quantum computers, yet their fault diagnosis is still dominated by threshold a

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

A Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent Post-Training

DGX agent

arXiv:2607.04574v1 Announce Type: cross Abstract: For LLM agents, supervised fine-tuning is not only about teacher labels' quality, but also about which interaction contexts those labels condition on.

safetyarxiv-cs-ai
7 Jul 2026
Model Releases

Agent-driven Long-tail Simulation for Autonomous Driving

DGX agent

arXiv:2607.04331v1 Announce Type: cross Abstract: Evaluating autonomous driving systems in closed-loop settings requires realistic and interactive simulation, yet existing simulators largely rely on l

model-releasesarxiv-cs-cv
7 Jul 2026
Model Releases

DrugAgent: Reliable Multi-Agent Integration of Conflicting Biomedical Evidence for Drug-Target Interaction Assessment

DGX agent

arXiv:2408.13378v5 Announce Type: replace Abstract: Workflows in drug-target interaction (DTI) assessment require integrating heterogeneous data from predictive models, curated resources, and observat

model-releasesarxiv-cs-ai
7 Jul 2026
Model Releases

GameEngineBench: Evaluating Coding Agents on Real C++ Runtime Environments

DGX agent

arXiv:2607.03525v1 Announce Type: cross Abstract: Game engines provide real-time simulation, rendering, physics, interaction, networking, and asset pipelines, making them valuable not only for games b

model-releasesarxiv-cs-cl
7 Jul 2026
Model Releases

PDEFlow: Autonomous Agentic PDE Pipelines for Neural Operator Learning and Solver-Free Inference

DGX agent

arXiv:2607.05134v1 Announce Type: cross Abstract: We present PDEFlow, an autonomous agentic framework that turns user-level ODE and PDE descriptions into solver-backed neural-operator pipelines. The w

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

Regime-Conditional Stabilisation of LLM-Augmented Cooperative Multi-Agent Reinforcement Learning

DGX agent

arXiv:2607.04470v1 Announce Type: cross Abstract: Large Language Models (LLMs) offer a natural interface for translating human objectives into reward signals for cooperative multi-agent reinforcement

safetyarxiv-cs-ai
7 Jul 2026
Hardware

SPORK: Self-Speculative Forking to Accelerate Agentic LLM Inference

DGX agent

arXiv:2607.03333v1 Announce Type: cross Abstract: LLM agents are becoming a common interface for research, coding, and question answering, yet their Thought-Action-Observation loop is often serial: th

hardwarearxiv-cs-ai
7 Jul 2026
Model Releases

A-TMA: Decoupling State-Aware Memory Failures in Long-Term Agent Memory

DGX agent

arXiv:2607.01935v1 Announce Type: new Abstract: Long term memory lets LLM agents act as persistent assistants, but user facts change. A useful memory system must know what is true now, what used to be

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

A^{2}utoLPBench: An Auto-Generated, Agent-Friendly LP Benchmark via Inverse-KKT Construction

DGX agent

arXiv:2607.02141v1 Announce Type: new Abstract: Most LP-from-text benchmarks are static datasets of word problems written and labeled by hand. Once such a dataset is released, its size is fixed, its d

model-releasesarxiv-cs-ai
3 Jul 2026
Local Ai

Auto-FL-Research: Agentic Search for Federated Learning Algorithms

DGX agent

arXiv:2607.01366v1 Announce Type: new Abstract: Federated learning (FL) research often depends on many small but consequential algorithmic choices: optimizer variants, server aggregation rules, local

local-aiarxiv-cs-ai
3 Jul 2026
Model Releases

Bringing Agentic Search to Earth Observation Data Discovery

DGX agent

arXiv:2607.02387v1 Announce Type: cross Abstract: NASA and its data centers hold thousands of geoscience datasets and tools like Worldview, Giovanni, the Science Discovery Engine, and Harmony. Finding

model-releasesarxiv-cs-lg
3 Jul 2026
Safety

Copewell: A Multi-Agent Swarm Architecture for Equitable Mental Wellness Support

DGX agent

arXiv:2607.02245v1 Announce Type: new Abstract: Mental health disorders affect nearly one billion people globally, yet 75% of individuals in low- and middle-income countries receive no treatment due t

safetyarxiv-cs-ai
3 Jul 2026
Local Ai

Repair the Amplifier, Not the Symptom: Stable World-Model Correction for Agent Rollouts

DGX agent

arXiv:2607.01767v1 Announce Type: new Abstract: As agent planning moves from short tool chains toward persistent workflows with thousands or tens of thousands of steps, failures will occur inside larg

local-aiarxiv-cs-ai
3 Jul 2026
Model Releases

Understanding Agent-Based Patching of Compiler Missed Optimizations

DGX agent

arXiv:2607.02370v1 Announce Type: cross Abstract: Compiler missed optimizations refer to cases in which compilers failed to optimize certain code. It takes many compiler developers' efforts to impleme

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

Multi-scale Mixture of World Models for Embodied Agents in Evolving Environments

DGX agent

arXiv:2607.00457v1 Announce Type: new Abstract: Embodied agents operating in the real world require multi-scale reasoning and knowledge adaptation as conditions change. We identify two challenges in a

safetyarxiv-cs-ai
2 Jul 2026
Safety

Self-Evolving Agents with Anytime-Valid Certificates

DGX agent

arXiv:2607.00871v1 Announce Type: new Abstract: Self-evolving agents violate the assumption behind most learning-theoretic guarantees: the data, evaluator, components, and hypothesis space are produce

safetyarxiv-cs-ai
2 Jul 2026
Hardware

AgRefactor: Self-Evolving Agentic Workflow for HLS Compatibility and Performance

DGX agent

arXiv:2606.30949v1 Announce Type: new Abstract: High-Level Synthesis (HLS) provides a fast path from concepts to silicon, but converting real-world software into synthesizable HLS code remains challen

hardwarearxiv-cs-ai
1 Jul 2026
Model Releases

AxDafny: Agentic Verified Code Generation in Dafny

DGX agent

arXiv:2606.32007v1 Announce Type: new Abstract: We study agentic code generation in Dafny, where a model must generate both executable code and the proof artifacts for verification. We present AxDafny

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

ECHO: Prune to act, trace to learn with selective turn memory in agentic RL

DGX agent

arXiv:2606.31650v1 Announce Type: cross Abstract: Long-horizon language agents must repeatedly interact with tools, accumulate evidence, and make decisions under bounded context windows. Existing cont

safetyarxiv-cs-ai
1 Jul 2026
Safety

IterCAD: An Iterative Multimodal Agent for Visually-Grounded CAD Generation and Editing

DGX agent

arXiv:2606.13368v2 Announce Type: replace Abstract: Computer-Aided Design is pivotal in modern manufacturing, yet existing automated methods predominantly rely on open-loop, one-shot generation, creat

safetyarxiv-cs-ai
1 Jul 2026
Safety

LabGuard: Grounding Natural-Language Laboratory Rules into Runtime Guards for Embodied Laboratory Agents

DGX agent

arXiv:2606.31045v1 Announce Type: new Abstract: Scientific embodied agents are increasingly capable of carrying out laboratory procedures, but executing these procedures safely in dynamic laboratory e

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

MECoBench: A Systematic Study of Multimodal Agent Collaboration in Embodied Environments

DGX agent

arXiv:2606.31966v1 Announce Type: cross Abstract: Recent multimodal large language models (MLLMs) have strong potential as embodied agents, but their ability to collaborate in visually grounded enviro

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

PPT-Eval: A Benchmark for Computer-Use Agents on PowerPoint Tasks

DGX agent

arXiv:2606.31154v1 Announce Type: cross Abstract: Creating and editing slides is a rich, multimodal activity that is ubiquitous in professional and educational settings, making it an ideal testbed for

model-releasesarxiv-cs-ai
1 Jul 2026
← Previous
1…979899100101…236
Next →