AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

Mean-Field Path-Integral Diffusion: From Samples to Interacting Agents

DGX agent

arXiv:2605.00007v1 Announce Type: cross Abstract: Independent sample generation is the prevailing paradigm in modern diffusion-based generative models of AI. We ask a different question: can samples c

model-releasesarxiv-cs-ai
5 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

MedScribe: Clinically Grounded CT Reporting through Agentic Workflows

DGX agent

arXiv:2605.01779v1 Announce Type: new Abstract: Vision-language models (VLMs) have shown potential for automated radiology report generation, yet existing approaches rely on global embedding compressi

agentsarxiv-cs-cv
5 May 2026
Agents

SiriusHelper: An LLM Agent-Based Operations Assistant for Big Data Platforms

DGX agent

arXiv:2605.00043v1 Announce Type: cross Abstract: Big data platforms are widely used in modern enterprises, and an in-production intelligent assistant is increasingly important to help users quickly f

agentsarxiv-cs-ai
5 May 2026
Model Releases

ExCyTIn-Bench: Evaluating LLM agents on Cyber Threat Investigation

DGX agent

arXiv:2507.14201v3 Announce Type: replace-cross Abstract: We present ExCyTIn-Bench, the first benchmark to Evaluate an LLM agent X on the task of Cyber Threat Investigation through security questions

model-releasesarxiv-cs-cl
4 May 2026
Model Releases

Agentic AI for Cybersecurity: A Meta-Cognitive Architecture for Governable Autonomy

DGX agent

arXiv:2602.11897v3 Announce Type: replace-cross Abstract: Cybersecurity decision-making increasingly occurs in environments characterized by uncertainty, partial observability, and adversarial manipul

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

AutoSurfer -- Teaching Web Agents through Comprehensive Surfing, Learning, and Modeling

DGX agent

arXiv:2604.27253v1 Announce Type: new Abstract: Recent advances in multimodal large language models (LLMs) have revolutionized web agents that can automate complex tasks on websites. However, their ac

model-releasesarxiv-cs-ai
1 May 2026
Safety

Bridging Values and Behavior: A Hierarchical Framework for Proactive Embodied Agents

DGX agent

arXiv:2604.27699v1 Announce Type: new Abstract: Current embodied agents are often limited to passive instruction-following or reactive need-satisfaction, lacking a stable, high-order value framework e

safetyarxiv-cs-ai
1 May 2026
Agents

Pragmos: A Process Agentic Modeling System

DGX agent

arXiv:2604.27311v1 Announce Type: cross Abstract: The advent of Large Language Models (LLMs) has significantly transformed tasks across Software Engineering. In the context of Business Process Managem

agentsarxiv-cs-ai
1 May 2026
Research

TiMem: Temporal-Hierarchical Memory Consolidation for Long-Horizon Conversational Agents

DGX agent

arXiv:2601.02845v2 Announce Type: replace-cross Abstract: Long-horizon conversational agents have to manage ever-growing interaction histories that quickly exceed the finite context windows of large l

researcharxiv-cs-ai
1 May 2026
Model Releases

A Systematic Comparison of Prompting and Multi-Agent Methods for LLM-based Stance Detection

DGX agent

arXiv:2604.26319v1 Announce Type: new Abstract: Stance detection identifies the attitude of a text author toward a given target. Recent studies have explored various LLM-based strategies for this task

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

AdaMem: Adaptive User-Centric Memory for Long-Horizon Dialogue Agents

DGX agent

arXiv:2603.16496v2 Announce Type: replace Abstract: Large language model (LLM) agents increasingly rely on external memory to support long-horizon interaction, personalized assistance, and multi-step

model-releasesarxiv-cs-cl
30 Apr 2026
Model Releases

Enforcing Benign Trajectories: A Behavioral Firewall for Structured-Workflow AI Agents

DGX agent

arXiv:2604.26274v1 Announce Type: cross Abstract: Structured-workflow agents driven by large language models execute tool calls against sensitive external environments. We propose odename, a telemetry

model-releasesarxiv-cs-ai
30 Apr 2026
Agents

RetroMotion: Retrocausal Motion Forecasting Models are Instructable

DGX agent

arXiv:2505.20414v2 Announce Type: replace-cross Abstract: Motion forecasts of road users (i.e., agents) vary in complexity depending on the number of agents, scene constraints, and interactions. In pa

agentsarxiv-cs-ai
30 Apr 2026
Local Ai

SecMate: Multi-Agent Adaptive Cybersecurity Troubleshooting with Tri-Context Personalization

DGX agent

arXiv:2604.26394v1 Announce Type: cross Abstract: Recent advances in large language models and agentic frameworks have enabled virtual customer assistants (VCAs) for complex support. We present SecMat

local-aiarxiv-cs-ai
30 Apr 2026
Agents

TDD Governance for Multi-Agent Code Generation via Prompt Engineering

DGX agent

arXiv:2604.26615v1 Announce Type: cross Abstract: Large language models (LLMs) accelerate software development but often exhibit instability, non-determinism, and weak adherence to development discipl

agentsarxiv-cs-ai
30 Apr 2026
Safety

Verified Critical Step Optimization for LLM Agents

DGX agent

arXiv:2602.03412v2 Announce Type: replace Abstract: As large language model agents tackle increasingly complex long-horizon tasks, effective post-training becomes critical. Prior work faces fundamenta

safetyarxiv-cs-cl
30 Apr 2026
Safety

Vibe Check: Understanding the Effects of LLM-Based Conversational Agents' Personality and Alignment on User Perceptions in Goal-Oriented Tasks

DGX agent

arXiv:2509.09870v2 Announce Type: replace-cross Abstract: Large language models (LLMs) enable conversational agents (CAs) to express distinctive personalities, raising new questions about how such des

safetyarxiv-cs-ai
30 Apr 2026
Safety

AdaRubric: Task-Adaptive Rubrics for LLM Agent Evaluation

DGX agent

arXiv:2603.21362v2 Announce Type: replace Abstract: LLM-as-Judge evaluation fails agent tasks because a fixed rubric cannot capture what matters for this task: code debugging demands Correctness and E

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

AgentHER: Hindsight Experience Replay for LLM Agent Trajectory Relabeling

DGX agent

arXiv:2603.21357v3 Announce Type: replace Abstract: LLM agents fail on the majority of real-world tasks -- GPT-4o succeeds on fewer than 15% of WebArena navigation tasks and below 55% pass@1 on ToolBe

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Beyond the Attention Stability Boundary: Agentic Self-Synthesizing Reasoning Protocols

DGX agent

arXiv:2604.24512v1 Announce Type: new Abstract: As LLM agents transition to autonomous digital coworkers, maintaining deterministic goal-directedness in non-linear multi-turn conversations emerged as

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

GCP: Guarded Collaborative Perception with Spatial-Temporal Aware Malicious Agent Detection

DGX agent

arXiv:2501.02450v2 Announce Type: replace Abstract: Collaborative perception significantly enhances autonomous driving safety by extending each vehicle's perception range through message sharing among

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

KLong: Training LLM Agent for Extremely Long-horizon Tasks

DGX agent

arXiv:2602.17547v3 Announce Type: replace Abstract: This paper introduces KLong, an open-source LLM agent trained to solve extremely long-horizon tasks. The principle is to first cold-start the model

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Peer Identity Bias in Multi-Agent LLM Evaluation: An Empirical Study Using the TRUST Democratic Discourse Analysis Pipeline

DGX agent

arXiv:2604.22971v1 Announce Type: cross Abstract: The TRUST democratic discourse analysis pipeline exposes its large language model (LLM) components to peer model identity through multiple structural

safetyarxiv-cs-ai
28 Apr 2026
Model Releases

PhysCodeBench: Benchmarking Physics-Aware Symbolic Simulation of 3D Scenes via Self-Corrective Multi-Agent Refinement

DGX agent

arXiv:2604.23580v1 Announce Type: cross Abstract: Physics-aware symbolic simulation of 3D scenes is critical for robotics, embodied AI, and scientific computing, requiring models to understand natural

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

TCOD: Exploring Temporal Curriculum in On-Policy Distillation for Multi-turn Autonomous Agents

DGX agent

arXiv:2604.24005v1 Announce Type: cross Abstract: On-policy distillation (OPD) has shown strong potential for transferring reasoning ability from frontier or domain-specific models to smaller students

safetyarxiv-cs-ai
28 Apr 2026
Agents

TeachMaster: Generative Teaching via Code

DGX agent

arXiv:2601.04204v2 Announce Type: replace-cross Abstract: The scalability of high-quality online education is hindered by the high costs and slow cycles of manual content creation. Despite advancement

agentsarxiv-cs-ai
28 Apr 2026
Agents

Vibe Medicine: Redefining Biomedical Research Through Human-AI Co-Work

DGX agent

arXiv:2604.23674v1 Announce Type: new Abstract: With the emergence of large language models (LLMs) and AI agent frameworks, the human-AI co-work paradigm known as Vibe Coding is changing how people co

agentsarxiv-cs-ai
28 Apr 2026
Applications

Your Reviews Replicate You: LLM-Based Agents as Customer Digital Twins for Conjoint Analysis

DGX agent

arXiv:2604.22756v1 Announce Type: cross Abstract: Conjoint analysis is a cornerstone of market research for estimating consumer preferences; however, traditional methods face persistent challenges reg

applicationsarxiv-cs-ai
28 Apr 2026
Research

Chain-of-Memory: Lightweight Memory Construction with Dynamic Evolution for LLM Agents

DGX agent

arXiv:2601.14287v2 Announce Type: replace Abstract: External memory systems are pivotal for enabling Large Language Model (LLM) agents to maintain persistent knowledge and perform long-horizon decisio

researcharxiv-cs-lg
27 Apr 2026
Agents

Fast Neural-Network Approximation of Active Target Search Under Uncertainty

DGX agent

arXiv:2604.22254v1 Announce Type: new Abstract: We address the problem of searching for an unknown number of stationary targets at unknown positions with a mobile agent. A probability hypothesis densi

agentsarxiv-cs-lg
27 Apr 2026
Model Releases

CI-Work: Benchmarking Contextual Integrity in Enterprise LLM Agents

DGX agent

arXiv:2604.21308v1 Announce Type: cross Abstract: Enterprise LLM agents can dramatically improve workplace productivity, but their core capability, retrieving and using internal context to act on a us

model-releasesarxiv-cs-cl
24 Apr 2026
Model Releases

Tool Attention Is All You Need: Dynamic Tool Gating and Lazy Schema Loading for Eliminating the MCP/Tools Tax in Scalable Agentic Workflows

DGX agent

arXiv:2604.21816v1 Announce Type: new Abstract: The Model Context Protocol (MCP) has become a common interface for connecting large language model (LLM) agents to external tools, but its reliance on s

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

SkillGraph: Graph Foundation Priors for LLM Agent Tool Sequence Recommendation

DGX agent

arXiv:2604.19793v1 Announce Type: new Abstract: LLM agents must select tools from large API libraries and order them correctly. Existing methods use semantic similarity for both retrieval and ordering

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

Supplement Generation Training for Enhancing Agentic Task Performance

DGX agent

arXiv:2604.20727v1 Announce Type: cross Abstract: Training large foundation models for agentic tasks is increasingly impractical due to the high computational costs, long iteration cycles, and rapid o

model-releasesarxiv-cs-ai
23 Apr 2026
Model Releases

A-MAR: Agent-based Multimodal Art Retrieval for Fine-Grained Artwork Understanding

DGX agent

arXiv:2604.19689v1 Announce Type: new Abstract: Understanding artworks requires multi-step reasoning over visual content and cultural, historical, and stylistic context. While recent multimodal large

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

M^{2}GRPO: Mamba-based Multi-Agent Group Relative Policy Optimization for Biomimetic Underwater Robots Pursuit

DGX agent

arXiv:2604.19404v1 Announce Type: cross Abstract: Traditional policy learning methods in cooperative pursuit face fundamental challenges in biomimetic underwater robots, where long-horizon decision ma

safetyarxiv-cs-ai
22 Apr 2026
Agents

Mind the (DH) Gap! A Contrast in Risky Choices Between Reasoning and Conversational LLMs

DGX agent

arXiv:2602.15173v2 Announce Type: replace Abstract: The use of large language models either as decision support systems, or in agentic workflows, is rapidly transforming the digital ecosystem. However

agentsarxiv-cs-ai
22 Apr 2026
Research

Sentipolis: Emotion-Aware Agents for Social Simulations

DGX agent

arXiv:2601.18027v2 Announce Type: replace Abstract: LLM agents are increasingly used for social simulation, yet emotion is often treated as a transient cue, causing emotional amnesia and weak long-hor

researcharxiv-cs-ai
22 Apr 2026
Safety

CAPO: Counterfactual Credit Assignment in Sequential Cooperative Teams

DGX agent

arXiv:2604.17693v1 Announce Type: new Abstract: In cooperative teams where agents act in a fixed order and share a single team reward, it is hard to know how much each agent contributed, and harder st

safetyarxiv-cs-lg
21 Apr 2026
Agents

Dual-Anchoring: Addressing State Drift in Vision-Language Navigation

DGX agent

arXiv:2604.17473v1 Announce Type: new Abstract: Vision-Language Navigation(VLN) requires an agent to navigate through 3D environments by following natural language instructions. While recent Video Lar

agentsarxiv-cs-cv
21 Apr 2026
Model Releases

Harness as an Asset: Enforcing Determinism via the Convergent AI Agent Framework (CAAF)

DGX agent

arXiv:2604.17025v1 Announce Type: cross Abstract: Large Language Models (LLMs) produce a controllability gap in safety-critical engineering: even low rates of undetected constraint violations render a

model-releasesarxiv-cs-lg
21 Apr 2026
Model Releases

Penny Wise, Pixel Foolish: Bypassing Price Constraints in Multimodal Agents via Visual Adversarial Perturbations

DGX agent

arXiv:2604.16515v1 Announce Type: new Abstract: The rapid proliferation of Multimodal Large Language Models (MLLMs) has enabled mobile agents to execute high-stakes financial transactions, but their a

model-releasesarxiv-cs-cv
21 Apr 2026
Agents

Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards

DGX agent

arXiv:2603.02008v2 Announce Type: replace Abstract: Effective exploration in reinforcement learning requires not only tracking where an agent has been, but also understanding how the agent perceives a

agentsarxiv-cs-lg
21 Apr 2026
Model Releases

MemEvoBench: Benchmarking Memory MisEvolution in LLM Agents

DGX agent

arXiv:2604.15774v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) with persistent memory enhances interaction continuity and personalization but introduces new safety risks. Speci

model-releasesarxiv-cs-cl
20 Apr 2026
Safety

Preregistered Belief Revision Contracts

DGX agent

arXiv:2604.15558v1 Announce Type: new Abstract: Deliberative multi-agent systems allow agents to exchange messages and revise beliefs over time. While this interaction is meant to improve performance,

safetyarxiv-cs-ai
20 Apr 2026
Agents

CCCE: A Continuous Code Calibration Engine for Autonomous Enterprise Codebase Maintenance via Knowledge Graph Traversal and Adaptive Decision Gating

DGX agent

arXiv:2604.13102v1 Announce Type: cross Abstract: Enterprise software organizations face an escalating challenge in maintaining the integrity, security, and freshness of codebases that span hundreds o

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

Don't Retrieve, Navigate: Distilling Enterprise Knowledge into Navigable Agent Skills for QA and RAG

DGX agent

arXiv:2604.14572v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) grounds LLM responses in external evidence but treats the model as a passive consumer of search results: it never

model-releasesarxiv-cs-cl
17 Apr 2026
Agents

EchoAgent: Towards Reliable Echocardiography Interpretation with 'Eyes','Hands' and 'Minds'

DGX agent

arXiv:2604.05541v2 Announce Type: replace Abstract: Reliable interpretation of echocardiography (Echo) is crucial for assessing cardiac function, which demands clinicians to synchronously orchestrate

agentsarxiv-cs-cv
17 Apr 2026
← Previous
1…9394959697…236
Next →