AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

LLM-Guided Agentic Floor Plan Parsing for Accessible Indoor Navigation of Blind and Low-Vision People

DGX agent

arXiv:2604.23970v1 Announce Type: new Abstract: Indoor navigation remains a critical accessibility challenge for the blind and low-vision (BLV) individuals, as existing solutions rely on costly per-bu

model-releasesarxiv-cs-ai
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

PExA: Parallel Exploration Agent for Complex Text-to-SQL

DGX agent

arXiv:2604.22934v1 Announce Type: new Abstract: LLM-based agents for text-to-SQL often struggle with latency-performance trade-off, where performance improvements come at the cost of latency or vice v

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

MolClaw: An Autonomous Agent with Hierarchical Skills for Drug Molecule Evaluation, Screening, and Optimization

DGX agent

arXiv:2604.21937v1 Announce Type: new Abstract: Computational drug discovery, particularly the complex workflows of drug molecule screening and optimization, requires orchestrating dozens of specializ

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Cross-Session Threats in AI Agents: Benchmark, Evaluation, and Algorithms

DGX agent

arXiv:2604.21131v1 Announce Type: cross Abstract: AI-agent guardrails are memoryless: each message is judged in isolation, so an adversary who spreads a single attack across dozens of sessions slips p

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Omission Constraints Decay While Commission Constraints Persist in Long-Context LLM Agents

DGX agent

arXiv:2604.20911v1 Announce Type: cross Abstract: LLM agents deployed in production operate under operator-defined behavioral policies (system-prompt instructions such as prohibitions on credential di

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

Reinforcement Learning with Foundation Priors: Let the Embodied Agent Efficiently Learn on Its Own

DGX agent

arXiv:2310.02635v5 Announce Type: replace-cross Abstract: Reinforcement learning (RL) is a promising approach for solving robotic manipulation tasks. However, it is challenging to apply the RL algorit

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Accelerating PayPal's Commerce Agent with Speculative Decoding: An Empirical Study on EAGLE3 with Fine-Tuned Nemotron Models

DGX agent

arXiv:2604.19767v1 Announce Type: cross Abstract: We evaluate speculative decoding with EAGLE3 as an inference-time optimization for PayPal's Commerce Agent, powered by a fine-tuned llama3.1-nemotron-

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Interval POMDP Shielding for Imperfect-Perception Agents

DGX agent

arXiv:2604.20728v1 Announce Type: new Abstract: Autonomous systems that rely on learned perception can make unsafe decisions when sensor readings are misclassified. We study shielding for this setting

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

SceneOrchestra: Efficient Agentic 3D Scene Synthesis via Full Tool-Call Trajectory Generation

DGX agent

arXiv:2604.19907v1 Announce Type: new Abstract: Recent agentic frameworks for 3D scene synthesis have advanced realism and diversity by integrating heterogeneous generation and editing tools. These to

model-releasesarxiv-cs-cv
23 Apr 2026
Safety

Semantic Prompting: Agentic Incremental Narrative Refinement through Spatial Semantic Interaction

DGX agent

arXiv:2604.19971v1 Announce Type: cross Abstract: Interactive spatial layouts empower users to synthesize information and organize findings for sensemaking. While Large Language Models (LLMs) can auto

safetyarxiv-cs-ai
23 Apr 2026
Safety

CentaurTA Studio: A Self-Improving Human-Agent Collaboration System for Thematic Analysis

DGX agent

arXiv:2604.18589v1 Announce Type: cross Abstract: Thematic analysis is difficult to scale: manual workflows are labor-intensive, while fully automated pipelines often lack controllability and transpar

safetyarxiv-cs-ai
22 Apr 2026
Safety

The Essence of Balance for Self-Improving Agents in Vision-and-Language Navigation

DGX agent

arXiv:2604.19064v1 Announce Type: new Abstract: In vision-and-language navigation (VLN), self-improvement from policy-induced experience, using only standard VLN action supervision, critically depends

safetyarxiv-cs-cv
22 Apr 2026
Model Releases

Visual Reasoning Agent: Robust Vision Systems in Remote Sensing via Inference-Time Scaling

DGX agent

arXiv:2509.16343v2 Announce Type: replace-cross Abstract: Building robust vision systems for high-stakes domains such as remote sensing requires stronger visual reasoning than what single-pass inferen

model-releasesarxiv-cs-ai
22 Apr 2026
Research

Beyond Verifiable Rewards: Rubric-Based GRM for Reinforced Fine-Tuning SWE Agents

DGX agent

arXiv:2604.16335v1 Announce Type: new Abstract: Despite recent progress in Large Language Model (LLM) Agents for Software Engineering (SWE) tasks, end-to-end fine-tuning typically relies on verifiable

researcharxiv-cs-lg
21 Apr 2026
Model Releases

BridgeEQA: Virtual Embodied Agents for Real Bridge Inspections

DGX agent

arXiv:2511.12676v2 Announce Type: replace Abstract: Deploying embodied agents that can answer questions about their surroundings in realistic real-world settings remains difficult, partly due to the s

model-releasesarxiv-cs-cv
21 Apr 2026
Model Releases

Evaluating Tool-Using Language Agents: Judge Reliability, Propagation Cascades, and Runtime Mitigation in AgentProp-Bench

DGX agent

arXiv:2604.16706v1 Announce Type: cross Abstract: Automated evaluation of tool-using large language model (LLM) agents is widely assumed to be reliable, but this assumption has rarely been validated a

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

HiGMem: A Hierarchical and LLM-Guided Memory System for Long-Term Conversational Agents

DGX agent

arXiv:2604.18349v1 Announce Type: new Abstract: Long-term conversational large language model (LLM) agents require memory systems that can recover relevant evidence from historical interactions withou

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Privacy-R1: Privacy-Aware Multi-LLM Agent Collaboration via Reinforcement Learning

DGX agent

arXiv:2510.16054v2 Announce Type: replace-cross Abstract: When users submit queries to Large Language Models (LLMs), their prompts can often contain sensitive data, forcing a difficult choice: Send th

local-aiarxiv-cs-cl
21 Apr 2026
Safety

Subliminal Transfer of Unsafe Behaviors in AI Agent Distillation

DGX agent

arXiv:2604.15559v1 Announce Type: new Abstract: Recent work on subliminal learning demonstrates that language models can transmit semantic traits through data that is semantically unrelated to those t

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

AccelOpt: A Self-Improving LLM Agentic System for AI Accelerator Kernel Optimization

DGX agent

arXiv:2511.15915v2 Announce Type: replace-cross Abstract: We present AccelOpt, a self-improving large language model (LLM) agentic system that autonomously optimizes kernels for emerging AI acclerator

model-releasesarxiv-cs-cl
17 Apr 2026
Model Releases

AgentIAD: Agentic Industrial Anomaly Detection via Adaptive Memory Augmentation

DGX agent

arXiv:2512.13671v2 Announce Type: replace Abstract: Industrial anomaly detection (IAD) is challenging due to the subtle and highly localized nature of many defects, which single-pass vision--language

model-releasesarxiv-cs-cv
17 Apr 2026
Model Releases

ContractSkill: Repairable Contract-Based Skills for Multimodal Web Agents

DGX agent

arXiv:2603.20340v3 Announce Type: replace-cross Abstract: Self-generated skills for web agents are often unstable and can even hurt performance relative to direct acting. We argue that the key bottlen

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

ECM Contracts: Contract-Aware, Versioned, and Governable Capability Interfaces for Embodied Agents

DGX agent

arXiv:2604.13097v1 Announce Type: cross Abstract: Embodied agents increasingly rely on modular capabilities that can be installed, upgraded, composed, and governed at runtime. Prior work has introduce

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

GeoAgentBench: A Dynamic Execution Benchmark for Tool-Augmented Agents in Spatial Analysis

DGX agent

arXiv:2604.13888v1 Announce Type: new Abstract: The integration of Large Language Models (LLMs) into Geographic Information Systems (GIS) marks a paradigm shift toward autonomous spatial analysis. How

model-releasesarxiv-cs-ai
17 Apr 2026
Model Releases

LLMOrbit: A Circular Taxonomy of Large Language Models -From Scaling Walls to Agentic AI Systems

DGX agent

arXiv:2601.14053v2 Announce Type: replace-cross Abstract: The field of artificial intelligence has undergone a revolution from foundational Transformer architectures to reasoning-capable systems appro

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

Evolvable Embodied Agent for Robotic Manipulation via Long Short-Term Reflection and Optimization

DGX agent

arXiv:2604.13533v1 Announce Type: cross Abstract: Achieving general-purpose robotics requires empowering robots to adapt and evolve based on their environment and feedback. Traditional methods face li

safetyarxiv-cs-cv
16 Apr 2026
Safety

Golden Handcuffs make safer AI agents

DGX agent

arXiv:2604.13609v1 Announce Type: new Abstract: Reinforcement learners can attain high reward through novel unintended strategies. We study a Bayesian mitigation for general environments: we expand th

safetyarxiv-cs-lg
16 Apr 2026
Applications

Memory Transfer Learning: How Memories are Transferred Across Domains in Coding Agents

DGX agent

arXiv:2604.14004v1 Announce Type: cross Abstract: Memory-based self-evolution has emerged as a promising paradigm for coding agents. However, existing approaches typically restrict memory utilization

applicationsarxiv-cs-cl
16 Apr 2026
Model Releases

MM-Doc-R1: Training Agents for Long Document Visual Question Answering through Multi-turn Reinforcement Learning

DGX agent

arXiv:2604.13579v1 Announce Type: new Abstract: Conventional Retrieval-Augmented Generation (RAG) systems often struggle with complex multi-hop queries over long documents due to their single-pass ret

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

Enhancing Agentic Textual Graph Retrieval with Synthetic Stepwise Supervision

DGX agent

arXiv:2510.03323v2 Announce Type: replace Abstract: Integrating textual graphs into Large Language Models (LLMs) is promising for complex graph-based QA. However, a key bottleneck is retrieving inform

safetyarxiv-cs-cl
15 Apr 2026
Model Releases

GCA Framework: A Gulf-Grounded Dataset and Agentic Pipeline for Climate Decision Support

DGX agent

arXiv:2604.12306v1 Announce Type: cross Abstract: Climate decision-making in the Gulf increasingly demands systems that can translate heterogeneous scientific and policy evidence into actionable guida

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Identity as Attractor: Geometric Evidence for Persistent Agent Architecture in LLM Activation Space

DGX agent

arXiv:2604.12016v1 Announce Type: new Abstract: Large language models map semantically related prompts to similar internal representations -- a phenomenon interpretable as attractor-like dynamics. We

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Self-Monitoring Benefits from Structural Integration: Lessons from Metacognition in Continuous-Time Multi-Timescale Agents

DGX agent

arXiv:2604.11914v1 Announce Type: new Abstract: Self-monitoring capabilities -- metacognition, self-prediction, and subjective duration -- are often proposed as useful additions to reinforcement learn

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

Agentic Application in Power Grid Static Analysis: Automatic Code Generation and Error Correction

DGX agent

arXiv:2604.09995v1 Announce Type: cross Abstract: This paper introduces an LLM agent that automates power grid static analysis by converting natural language into MATPOWER scripts. The framework utili

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

AutoMS: Multi-Agent Evolutionary Search for Cross-Physics Inverse Microstructure Design

DGX agent

arXiv:2603.27195v2 Announce Type: replace Abstract: Designing microstructures with coupled cross-physics objectives is a fundamental challenge where traditional topology optimization is often computat

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

BLUEmed: Retrieval-Augmented Multi-Agent Debate for Clinical Error Detection

DGX agent

arXiv:2604.10389v1 Announce Type: new Abstract: Terminology substitution errors in clinical notes, where one medical term is replaced by a linguistically valid but clinically different term, pose a pe

model-releasesarxiv-cs-cl
14 Apr 2026
Tutorials

Controlling Multimodal Conversational Agents with Coverage-Enhanced Latent Actions

DGX agent

arXiv:2601.07516v2 Announce Type: replace-cross Abstract: Vision-language models are increasingly employed as multimodal conversational agents (MCAs) for diverse conversational tasks. Recently, reinfo

tutorialsarxiv-cs-ai
14 Apr 2026
Model Releases

DERM-3R: A Resource-Efficient Multimodal Agents Framework for Dermatologic Diagnosis and Treatment in Real-World Clinical Settings

DGX agent

arXiv:2604.09596v1 Announce Type: new Abstract: Dermatologic diseases impose a large and growing global burden, affecting billions and substantially reducing quality of life. While modern therapies ca

model-releasesarxiv-cs-ai
14 Apr 2026
Safety

Dialectic-Med: Mitigating Diagnostic Hallucinations via Counterfactual Adversarial Multi-Agent Debate

DGX agent

arXiv:2604.11258v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) in healthcare suffer from severe confirmation bias, often hallucinating visual details to support initial, pote

safetyarxiv-cs-cl
14 Apr 2026
Safety

EvoNash-MARL: A Closed-Loop Multi-Agent Reinforcement Learning Framework for Medium-Horizon Equity Allocation

DGX agent

arXiv:2604.10911v1 Announce Type: new Abstract: Medium-to-long-horizon stock allocation presents significant challenges due toveak predictive structures, non-stadonary market regimes, and the degradat

safetyarxiv-cs-ai
14 Apr 2026
Model Releases

RCBSF: A Multi-Agent Framework for Automated Contract Revision via Stackelberg Game

DGX agent

arXiv:2604.10740v1 Announce Type: new Abstract: Despite the widespread adoption of Large Language Models (LLMs) in Legal AI, their utility for automated contract revision remains impeded by hallucinat

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

RISK: A Framework for GUI Agents in E-commerce Risk Management

DGX agent

arXiv:2509.21982v2 Announce Type: replace Abstract: E-commerce risk management requires aggregating diverse, deeply embedded web data through multi-step, stateful interactions, which traditional scrap

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

SCMAPR: Self-Correcting Multi-Agent Prompt Refinement for Complex-Scenario Text-to-Video Generation

DGX agent

arXiv:2604.05489v3 Announce Type: replace Abstract: Text-to-Video (T2V) generation has benefited from recent advances in diffusion models, yet current systems still struggle under complex scenarios, w

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

The Missing Knowledge Layer in Cognitive Architectures for AI Agents

DGX agent

arXiv:2604.11364v1 Announce Type: new Abstract: The two most influential cognitive architecture frameworks for AI agents, CoALA [21] and JEPA [12], both lack an explicit Knowledge layer with its own p

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

UniToolCall: Unifying Tool-Use Representation, Data, and Evaluation for LLM Agents

DGX agent

arXiv:2604.11557v1 Announce Type: new Abstract: Tool-use capability is a fundamental component of LLM agents, enabling them to interact with external systems through structured function calls. However

model-releasesarxiv-cs-ai
14 Apr 2026
Model Releases

ZARA: Training-Free Motion Time-Series Reasoning via Evidence-Grounded LLM Agents

DGX agent

arXiv:2508.04038v2 Announce Type: replace Abstract: Motion sensor time-series are central to Human Activity Recognition (HAR), yet conventional approaches are constrained to fixed activity sets and ty

model-releasesarxiv-cs-cl
14 Apr 2026
Model Releases

AlphaLab: Autonomous Multi-Agent Research Across Optimization Domains with Frontier LLMs

DGX agent

arXiv:2604.08590v1 Announce Type: cross Abstract: We present AlphaLab, an autonomous research harness that leverages frontier LLM agentic capabilities to automate the full experimental cycle in quanti

model-releasesarxiv-cs-ai
13 Apr 2026
Model Releases

DRBENCHER: Can Your Agent Identify the Entity, Retrieve Its Properties and Do the Math?

DGX agent

arXiv:2604.09251v1 Announce Type: new Abstract: Deep research agents increasingly interleave web browsing with multi-step computation, yet existing benchmarks evaluate these capabilities in isolation,

model-releasesarxiv-cs-ai
13 Apr 2026
← Previous
1…103104105106107…236
Next →