AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Safety

TAPO: Transition-Aware Policy Optimization for LLM Agents

DGX agent

arXiv:2607.27973v1 Announce Type: new Abstract: Recently, Reinforcement Learning (RL) has emerged as a crucial paradigm for the post-training of Large Language Model (LLM) agents. However, existing me

safetyarxiv-cs-lg
31 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Agentic AI in medicine: architectures, applications, evaluation, and challenges for clinical translation

DGX agent

arXiv:2607.25489v1 Announce Type: new Abstract: Large language models and multimodal foundation models are enabling medical artificial intelligence (AI) systems to move beyond isolated prediction and

safetyarxiv-cs-cv
29 Jul 2026
Model Releases

Cooperative Multi-UAV Navigation in Complex Environments via Systematic Multi-Agent Deep Reinforcement Learning

DGX agent

arXiv:2607.25754v1 Announce Type: new Abstract: Cooperative navigation of multi-agent UAVs in complex environments faces key challenges including local optima traps, sparse rewards, learning imbalance

model-releasesarxiv-cs-ro
29 Jul 2026
Model Releases

PATHFinder Agent for Tailored Prenatal Care

DGX agent

arXiv:2607.24768v1 Announce Type: new Abstract: Prenatal care is an important preventive service designed to improve outcomes for pregnant individuals. The American College of Obstetricians and Gyneco

model-releasesarxiv-cs-ai
29 Jul 2026
Local Ai

ProcAgent: An Agentic Framework for Procedural Task Guidance on Edge with Human-in-the-Loop

DGX agent

arXiv:2607.24770v1 Announce Type: new Abstract: Procedural tasks such as furniture assembly and home repair impose substantial cognitive demands because users must interpret instructions, track task p

local-aiarxiv-cs-ai
29 Jul 2026
Model Releases

RSMeM: Knowledge-Enhanced Memory Evolution for Remote Sensing Agents with Systematic Evaluation

DGX agent

arXiv:2607.24772v1 Announce Type: new Abstract: Geoscience research requires complex analysis and domain expertise, with remote sensing (RS) observations as a key foundation. However, existing RS agen

model-releasesarxiv-cs-ai
29 Jul 2026
Model Releases

Evaluating Fuzz Testing for Reinforcement Learning Agents

DGX agent

arXiv:2607.24577v1 Announce Type: new Abstract: Reinforcement Learning (RL) agents are increasingly deployed in safety-critical domains such as robotics, autonomous driving, and drone control, where u

model-releasesarxiv-cs-lg
28 Jul 2026
Model Releases

Gubernaut: A Deterministic Homeostatic Controller for Affect-Regulated LLM Agents, Validated Across Independent Model Families

DGX agent

arXiv:2607.24339v1 Announce Type: new Abstract: Large language model (LLM) agents inherit reactive failure modes: escalation under provocation, sycophantic drift under flattery, perseveration when stu

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

HELIOS: An LLM-Driven Autonomous Indirect Trajectory Optimization Agent

DGX agent

arXiv:2607.24051v1 Announce Type: cross Abstract: Low-thrust trajectory optimization is a core technology in deep-space mission design. Indirect methods based on Pontryagin's Minimum Principle (PMP) o

agentsarxiv-cs-ai
28 Jul 2026
Model Releases

LazyMem: Retrieve Broadly, Construct Selectively for Efficient Long-Term Agent Memory

DGX agent

arXiv:2607.22690v1 Announce Type: new Abstract: Long-term memory lets LLM agents reuse past interactions, but raw dialogue histories are verbose and information-sparse. Retrieving broadly improves evi

model-releasesarxiv-cs-ai
28 Jul 2026
Safety

Looping Is Not Reliability: State-Bound Evidence and Typed Revision Contracts for Agentic Code Repair

DGX agent

arXiv:2607.24604v1 Announce Type: cross Abstract: Generate--test--revise loops are common in coding agents, but repetition alone provides no reliability guarantee. We study the gap between finding a c

safetyarxiv-cs-ai
28 Jul 2026
Safety

MedDDC-Eval: Diagnosis-Decoupled Evaluation of Multi-Turn Medical Consultation Agents

DGX agent

arXiv:2607.18999v2 Announce Type: replace-cross Abstract: Evaluating multi-turn medical consultation agents requires judging the diagnostic support provided by the histories they elicit through intera

safetyarxiv-cs-ai
28 Jul 2026
Model Releases

QFoldAgent: An Autonomous Quantum Optimization Multi-Agent System for Protein Structure Prediction

DGX agent

arXiv:2607.22549v1 Announce Type: new Abstract: Hybrid quantum-classical protein structure prediction depends strongly on Hamiltonian penalty weights, yet existing lattice-based workflows typically fi

model-releasesarxiv-cs-ai
28 Jul 2026
Agents

Agentic Root Cause Analysis through Evidence-Grounded Reasoning

DGX agent

arXiv:2607.22385v1 Announce Type: cross Abstract: Diagnosing the root cause of anomalies is essential for safe industrial operation. Despite extensive sensor instrumentation, formulating hypotheses an

agentsarxiv-cs-lg
27 Jul 2026
Safety

Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning

DGX agent

arXiv:2607.21653v1 Announce Type: cross Abstract: Agentic reinforcement learning research is constant algorithm modification, new estimators, new pipeline stages, new rollout schemes, and in mainstrea

safetyarxiv-cs-cl
27 Jul 2026
Agents

Agent-Guided Relational Concept Discovery: Toward Interpretable Surgical Margin Assessment

DGX agent

arXiv:2607.21437v1 Announce Type: new Abstract: Deep learning models can effectively use Rapid Evaporative Ionization Mass Spectrometry (REIMS) data for surgical margin assessment. However, their clin

agentsarxiv-cs-ai
24 Jul 2026
Safety

ARCO: Adaptive Rubrics with Co-Evolution for Multi-Step LLM-Based Agents

DGX agent

arXiv:2606.21262v2 Announce Type: replace Abstract: Reinforcement learning for multi-step LLM agents often relies on scalar rewards that indicate success but cannot explain why a trajectory is good or

safetyarxiv-cs-ai
24 Jul 2026
Model Releases

AREX: Towards a Recursively Self-Improving Agent for Deep Research

DGX agent

arXiv:2607.21461v1 Announce Type: new Abstract: Deep research requires agents to find answers that jointly satisfy multiple constraints. Discovering such answers is costly, whereas verifying a candida

model-releasesarxiv-cs-ai
24 Jul 2026
Agents

Bayesian uncertainty estimation improves clinical decision making in medical AI agents

DGX agent

arXiv:2607.20582v1 Announce Type: cross Abstract: Machine learning models for medical image analysis typically lack a reliable measure of confidence, limiting their use in ambiguous or atypical cases.

agentsarxiv-cs-ai
24 Jul 2026
Agents

OPTScientist: Multi-Agent Discovery of Typed Optimizer Programs for Transformer Pretraining

DGX agent

arXiv:2607.20486v1 Announce Type: new Abstract: Designing optimizers for modern deep learning remains a challenging scientific problem, requiring the joint consideration of optimization geometry, stat

agentsarxiv-cs-ai
24 Jul 2026
Safety

The Dark Room in the Reward Channel: Dense Prediction Rewards Collapse GRPO-Trained LLM Agents -- and What Actually Works

DGX agent

arXiv:2607.21273v1 Announce Type: new Abstract: Dense per-step supervision is an appealing remedy for sparse-reward, long-horizon LLM agents: reward the agent for predicting its next observation, and

safetyarxiv-cs-lg
24 Jul 2026
Model Releases

Workflow-Localized Mechanism Learning: Attribution-Guided Repair and Knowledge Reuse for Structured Agent Skills

DGX agent

arXiv:2607.20999v1 Announce Type: new Abstract: Agent Skills package reusable procedural knowledge as external artifacts for frozen language-model agents, yet existing optimizers do not jointly resolv

model-releasesarxiv-cs-ai
24 Jul 2026
Model Releases

AgentCgroup: Understanding and Controlling OS Resources of AI Agents

DGX agent

arXiv:2602.09345v3 Announce Type: replace-cross Abstract: AI agents are increasingly deployed in multi-tenant cloud environments, where they execute diverse tool calls within sandboxed containers, eac

model-releasesarxiv-cs-ai
23 Jul 2026
Agents

An LLM-powered Agentic Recommendation System for Connected TV Content Discovery

DGX agent

arXiv:2607.09988v3 Announce Type: replace-cross Abstract: Recommendation systems, from traditional multi-stage to recent unified generative architectures, face challenges in incorporating diverse cont

agentsarxiv-cs-ai
23 Jul 2026
Model Releases

CEO-Bench: Can Agents Play the Long Game?

DGX agent

arXiv:2606.18543v2 Announce Type: replace Abstract: Language model agents are becoming proficient executors at isolated, short-horizon tasks such as software engineering and customer service. Yet real

model-releasesarxiv-cs-ai
23 Jul 2026
Safety

Coordinating from Memory: Graph-Structured Experience Reuse for Multi-Agent Adaptation in Dynamic Manufacturing

DGX agent

arXiv:2607.19985v1 Announce Type: new Abstract: Dynamic manufacturing environments require multi-agent systems to coordinate effectively under frequent operational disturbances such as machine failure

safetyarxiv-cs-ai
23 Jul 2026
Safety

Guardrails as Scapegoats: Auditing Unfaithful Safety Refusals in Tool-Augmented LLM Agents

DGX agent

arXiv:2607.19449v1 Announce Type: cross Abstract: Evaluation frameworks for tool-augmented LLM agents focus overwhelmingly on capability metrics or explicit tool crashes, leaving silent infrastructure

safetyarxiv-cs-ai
23 Jul 2026
Model Releases

The Chronos Vulnerability: A Taxonomy of Temporal Persistence and Memory-Based Deception in Agentic AI

DGX agent

arXiv:2607.19433v1 Announce Type: new Abstract: The transition from stateless generative models in artificial intelligence to stateful, autonomous agents represents an architectural evolution that, wh

model-releasesarxiv-cs-ai
23 Jul 2026
Model Releases

COLMAR: Cooperative View Policy Learning for Multi-Agent Active 3D Reconstruction

DGX agent

arXiv:2607.13524v1 Announce Type: new Abstract: Active 3D reconstruction requires selecting informative viewpoints under limited sensing budgets. In multi-agent settings, coordination inefficiencies s

model-releasesarxiv-cs-ro
16 Jul 2026
Model Releases

Exploratory, Communicative, and Deployable: Vision-Driven Embodied Agents for Open-World Mobile Manipulation

DGX agent

arXiv:2607.13653v1 Announce Type: new Abstract: Real-world deployment of embodied agents requires active exploration, visual grounding, and interactive intent disambiguation. However, existing framewo

model-releasesarxiv-cs-cv
16 Jul 2026
Model Releases

Inference Economics of Enterprise Coding Agents: A Case Study of Cloud vs. On-Premise LLMs

DGX agent

arXiv:2607.13080v1 Announce Type: cross Abstract: Autonomous coding agents force engineering organizations to choose between API-based frontier models -- strong reasoning at high token cost -- and on-

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Learning Engagement Assistant (LEA): Cross-Course Scalability and Classroom Evaluation of an Agentic AI Tutoring System

DGX agent

arXiv:2607.13370v1 Announce Type: cross Abstract: This paper is an extension of a paper presented at the ICAART 2026 conference, which introduced LEA (Learning Engagement Assistant), an adaptive AI tu

agentsarxiv-cs-ai
16 Jul 2026
Safety

Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents

DGX agent

arXiv:2607.13591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly rely on external memory systems to accumulate experience across tasks. Yet nearly all existing approach

safetyarxiv-cs-ai
16 Jul 2026
Model Releases

TRACE: Turn-level Reward Assignment via Credit Estimation for Long-Horizon Agents

DGX agent

arXiv:2607.13988v1 Announce Type: new Abstract: Multi-turn agents solve complex tasks through extended sequences of tool interactions before producing a final answer, making credit assignment a fundam

model-releasesarxiv-cs-lg
16 Jul 2026
Agents

Game Theory Driven Multi-Agent Framework Mitigates Language Model Hallucination

DGX agent

arXiv:2607.08403v1 Announce Type: new Abstract: The application of lightweight Large Language Models in rule-based scientific domains remains severely limited by their tendency to mimic linguistic pat

agentsarxiv-cs-ai
10 Jul 2026
Model Releases

RetailBench: Evaluating Long-Horizon Autonomous Decision-Making and Strategy Stability of LLM Agents in Realistic Retail Environments

DGX agent

arXiv:2603.16453v3 Announce Type: replace Abstract: Large language model (LLM) agents have made rapid progress on short-horizon, well-scoped tasks, yet their ability to sustain coherent decisions in d

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

What to Keep, What to Forget: A Rate--Distortion View of Memory Compaction in LLMs and Agents

DGX agent

arXiv:2607.08032v1 Announce Type: new Abstract: Large language models, and the agents built on them, spend an ever-growing share of their compute and memory on remembering: caching attention keys and

model-releasesarxiv-cs-lg
10 Jul 2026
Model Releases

A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong

DGX agent

arXiv:2607.06854v1 Announce Type: cross Abstract: Reinforcement learning agents for imperfect-information card games are only as strong as the opponents they train against, and they are hard to grade,

model-releasesarxiv-cs-ai
9 Jul 2026
Model Releases

Does AI Understand Imaging? A Systematic Benchmark of Agentic AI for Computational Imaging Tasks

DGX agent

arXiv:2607.07189v1 Announce Type: new Abstract: Vision-language models (VLMs) and agentic AI have shown strong performance on semantic visual tasks, but it remains unclear whether they can handle the

model-releasesarxiv-cs-ai
9 Jul 2026
Agents

End-to-End LLM Flight Planning with RAG-based Memory and Multi-modal Coach Agent

DGX agent

arXiv:2607.06964v1 Announce Type: cross Abstract: Bridging the gap between human pilot intent and autonomous flight operation is critical for real-world electric vertical takeoff and landing (eVTOL) a

agentsarxiv-cs-ai
9 Jul 2026
Agents

Reliable and Developer-Aligned Evaluation of Agents for Software Engineering

DGX agent

arXiv:2607.06713v1 Announce Type: cross Abstract: Large language models are rapidly moving towards closing the development cycle, transitioning from simple assistive companions to autonomous contribut

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

Beyond Reactivity: Measuring Proactive Problem Solving in LLM Agents

DGX agent

arXiv:2510.19771v4 Announce Type: replace Abstract: LLM-based agents are increasingly moving towards proactivity: rather than awaiting instruction, they exercise agency to anticipate user needs and so

model-releasesarxiv-cs-ai
8 Jul 2026
Local Ai

CHARLIE: An On-Premise Multi-Agent Retrieval-Augmented Generation System for Evidential Reasoning in Forensic Science

DGX agent

arXiv:2607.05428v1 Announce Type: cross Abstract: We present Charlie, an on-premise multi-agent Retrieval-Augmented Generation (RAG) system for structured evidential processing in digital forensic env

local-aiarxiv-cs-ai
8 Jul 2026
Model Releases

Doomed from the Start: Early Abort of LLM Agent Episodes via a Recall-Controlled Probe Cascade

DGX agent

arXiv:2607.06503v1 Announce Type: new Abstract: Large language model (LLM) agents solving multi-step tasks frequently commit to trajectories that are doomed to fail, yet continue to consume substantia

model-releasesarxiv-cs-ai
8 Jul 2026
Agents

From Voting to Agent Collaboration: Answer-Type-Aware LLM Pipelines for BioASQ 14b

DGX agent

arXiv:2607.06452v1 Announce Type: cross Abstract: Biomedical question answering requires not only accurate extraction of information from scientific literature but also reliable integration of evidenc

agentsarxiv-cs-ai
8 Jul 2026
Hardware

Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory

DGX agent

arXiv:2607.05511v1 Announce Type: new Abstract: Agentic video understanding equips models with long-term memory to autonomously process and respond to continuous, long-horizon multimodal streams. Howe

hardwarearxiv-cs-cv
8 Jul 2026
Agents

Multi-Agent Deep Reinforcement Learning for Multi Objective Battery Management in Dairy Farms

DGX agent

arXiv:2607.06489v1 Announce Type: new Abstract: The dairy industry in Ireland has a large potential for the integration of renewable energy and the reduction of carbon emissions. However, researchers

agentsarxiv-cs-ai
8 Jul 2026
Agents

PolyJarvis: An LLM-Orchestrated Agent for Automated All-Atom Molecular Dynamics of Amorphous Homopolymers

DGX agent

arXiv:2604.02537v2 Announce Type: replace Abstract: All-atom molecular dynamics (MD) simulations can predict polymer properties from molecular structure, yet their execution requires specialized exper

agentsarxiv-cs-cl
8 Jul 2026
← Previous
1…6566676869…236
Next →