AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Safety

The Endogeneity of Miscalibration: Impossibility and Escape in Scored Reporting

DGX agent

arXiv:2605.07671v1 Announce Type: cross Abstract: Eliciting truthful reports from autonomous agents is a core problem in scalable AI oversight: a principal scores the agent's report using a strictly p

safetyarxiv-cs-ai
11 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ProgramBench: Can Language Models Rebuild Programs From Scratch?

DGX agent

arXiv:2605.03546v1 Announce Type: cross Abstract: Turning ideas into full software projects from scratch has become a popular use case for language models. Agents are being deployed to seed, maintain,

agentsarxiv-cs-ai
7 May 2026
Agents

SPHERE: Mitigating the Loss of Spectral Plasticity in Mixture-of-Experts for Deep Reinforcement Learning

DGX agent

arXiv:2605.04712v1 Announce Type: new Abstract: In deep reinforcement learning (DRL), an agent is trained from a stream of experience. In a continual learning setting, such agents can suffer from plas

agentsarxiv-cs-lg
7 May 2026
Model Releases

DataEvolver: Let Your Data Build and Improve Itself via Goal-Driven Loop Agents

DGX agent

arXiv:2605.01789v1 Announce Type: new Abstract: Constructing controllable visual data is a major bottleneck for image editing and multimodal understanding. Useful supervision is rarely produced by a s

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

LLM-ADAM: A Generalizable LLM Agent Framework for Pre-Print Anomaly Detection in Additive Manufacturing

DGX agent

arXiv:2605.03328v1 Announce Type: new Abstract: Additive manufacturing (AM) continues to transform modern manufacturing by enabling flexible, on-demand production of complex geometries across diverse

model-releasesarxiv-cs-lg
6 May 2026
Agents

Separating Intelligence from Execution: A Workflow Engine for the Model Context Protocol

DGX agent

arXiv:2605.00827v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly interact with external systems through tool-calling protocols such as the Model Context Protocol (MCP).

agentsarxiv-cs-ai
6 May 2026
Safety

T^2PO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement Learning

DGX agent

arXiv:2605.02178v1 Announce Type: new Abstract: Recent progress in multi-turn reinforcement learning (RL) has significantly improved reasoning LLMs' performances on complex interactive tasks. Despite

safetyarxiv-cs-ai
6 May 2026
Agents

Networked Information Aggregation for Binary Classification

DGX agent

arXiv:2605.01082v1 Announce Type: new Abstract: We study networked binary classification on a directed acyclic graph (DAG) where each agent observes only a subset of the feature columns of a shared da

agentsarxiv-cs-lg
5 May 2026
Applications

Autoformalizing Memory Specifications with Agents

DGX agent

arXiv:2605.00058v1 Announce Type: cross Abstract: The primary goal of Design Verification (DV) is to ensure that a proposed chip design implementation (either in code, or physical form) exactly matche

applicationsarxiv-cs-lg
4 May 2026
Local Ai

Enhancing Linux Privilege Escalation Attack Capabilities of Local LLM Agents

DGX agent

arXiv:2604.27143v1 Announce Type: cross Abstract: Recent research has demonstrated the potential of Large Language Models (LLMs) for autonomous penetration testing, particularly when using cloud-based

local-aiarxiv-cs-ai
1 May 2026
Research

Learning to Aggregate Zero-Shot LLM Agents for Corporate Disclosure Classification

DGX agent

arXiv:2603.20965v2 Announce Type: replace-cross Abstract: This paper studies whether a lightweight supervised aggregator can combine diverse zero-shot large language model outputs into a stronger down

researcharxiv-cs-ai
1 May 2026
Agents

Synthetic Computers at Scale for Long-Horizon Productivity Simulation

DGX agent

arXiv:2604.28181v1 Announce Type: new Abstract: Realistic long-horizon productivity work is strongly conditioned on user-specific computer environments, where much of the work context is stored and or

agentsarxiv-cs-ai
1 May 2026
Model Releases

Learning to Ask: When LLM Agents Meet Unclear Instruction

DGX agent

arXiv:2409.00557v4 Announce Type: replace-cross Abstract: Equipped with the capability to call functions, modern large language models (LLMs) can leverage external tools for addressing a range of task

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis

DGX agent

arXiv:2604.16922v2 Announce Type: replace Abstract: Climate research is pivotal for mitigating global environmental crises, yet the accelerating volume of multi-scale datasets and the complexity of an

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

SeaEvo: Advancing Algorithm Discovery with Strategy Space Evolution

DGX agent

arXiv:2604.24372v1 Announce Type: cross Abstract: LLM-guided evolutionary search has emerged as a promising paradigm for automated algorithm discovery, yet most systems track search progress primarily

agentsarxiv-cs-ai
28 Apr 2026
Agents

On the Hybrid Nature of ABPMS Process Frames and its Implications on Automated Process Discovery

DGX agent

arXiv:2604.22455v1 Announce Type: new Abstract: A core component of any AI-Augmented Business Process Management System (ABPMS) is the process frame, which gives the system process-awareness and defin

agentsarxiv-cs-ai
27 Apr 2026
Model Releases

SlideAgent: Hierarchical Agentic Framework for Multi-Page Visual Document Understanding

DGX agent

arXiv:2510.26615v3 Announce Type: replace Abstract: Retrieval-augmented generation (RAG) extends large language models (LLMs) with external knowledge, but it must balance limited effective context, re

model-releasesarxiv-cs-cl
24 Apr 2026
Safety

Diagnosing CFG Interpretation in LLMs

DGX agent

arXiv:2604.20811v1 Announce Type: new Abstract: As LLMs are increasingly integrated into agentic systems, they must adhere to dynamically defined, machine-interpretable interfaces. We evaluate LLMs as

safetyarxiv-cs-ai
23 Apr 2026
Agents

Self-Guided Plan Extraction for Instruction-Following Tasks with Goal-Conditional Reinforcement Learning

DGX agent

arXiv:2604.20601v1 Announce Type: new Abstract: We introduce SuperIgor, a framework for instruction-following tasks. Unlike prior methods that rely on predefined subtasks, SuperIgor enables a language

agentsarxiv-cs-ai
23 Apr 2026
Safety

Large Language Models Exhibit Normative Conformity

DGX agent

arXiv:2604.19301v1 Announce Type: new Abstract: The conformity bias exhibited by large language models (LLMs) can pose a significant challenge to decision-making in LLM-based multi-agent systems (LLM-

safetyarxiv-cs-ai
22 Apr 2026
Agents

CogDriver: Integrating Cognitive Inertia for Temporally Coherent Planning in Autonomous Driving

DGX agent

arXiv:2509.00789v2 Announce Type: replace Abstract: The pursuit of autonomous agents capable of temporally coherent planning is hindered by a fundamental flaw in current vision-language models (VLMs):

agentsarxiv-cs-cv
21 Apr 2026
Research

EmbodiedHead: Real-Time Listening and Speaking Avatar for Conversational Agents

DGX agent

arXiv:2604.17211v1 Announce Type: new Abstract: We present EmbodiedHead, a speech-driven talking-head framework that equips LLMs with real-time visual avatars for conversation. A practical embodied av

researcharxiv-cs-cv
21 Apr 2026
Safety

Empowering Multi-Turn Tool-Integrated Agentic Reasoning with Group Turn Policy Optimization

DGX agent

arXiv:2511.14846v2 Announce Type: replace-cross Abstract: Training Large Language Models (LLMs) for multi-turn Tool-Integrated Reasoning (TIR) - where models iteratively reason, generate code, and ver

safetyarxiv-cs-cl
21 Apr 2026
Research

From Verbatim to Gist: Distilling Pyramidal Multimodal Memory via Semantic Information Bottleneck for Long-Horizon Video Agents

DGX agent

arXiv:2603.01455v2 Announce Type: replace-cross Abstract: While multimodal large language models have demonstrated impressive short-term reasoning, they struggle with long-horizon video understanding

researcharxiv-cs-cl
21 Apr 2026
Model Releases

Latent Preference Modeling for Cross-Session Personalized Tool Calling

DGX agent

arXiv:2604.17886v1 Announce Type: new Abstract: Users often omit essential details in their requests to LLM-based agents, resulting in under-specified inputs for tool use. This poses a fundamental cha

model-releasesarxiv-cs-cl
21 Apr 2026
Research

TSegAgent: Zero-Shot Tooth Segmentation via Geometry-Aware Vision-Language Agents

DGX agent

arXiv:2603.19684v2 Announce Type: replace Abstract: Automatic tooth segmentation and identification from intra-oral scanned 3D models are fundamental problems in digital dentistry, yet most existing a

researcharxiv-cs-cv
21 Apr 2026
Agents

Optimal Solutions for the Moving Target Vehicle Routing Problem via Branch-and-Price with Relaxed Continuity

DGX agent

arXiv:2603.00663v4 Announce Type: replace Abstract: The Moving Target Vehicle Routing Problem (MT-VRP) seeks trajectories for several agents that intercept a set of moving targets, subject to speed, t

agentsarxiv-cs-ro
20 Apr 2026
Agents

SENSE: Stereo OpEN Vocabulary SEmantic Segmentation

DGX agent

arXiv:2604.15946v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation enables models to segment objects or image regions beyond fixed class sets, offering flexibility in dynamic enviro

agentsarxiv-cs-cv
20 Apr 2026
Agents

An Intelligent Robotic and Bio-Digestor Framework for Smart Waste Management

DGX agent

arXiv:2604.14882v1 Announce Type: cross Abstract: Rapid urbanization and continuous population growth have made municipal solid waste management increasingly challenging. These challenges highlight th

agentsarxiv-cs-lg
17 Apr 2026
Model Releases

Applying an Agentic Coding Tool for Improving Published Algorithm Implementations

DGX agent

arXiv:2604.13109v1 Announce Type: cross Abstract: We present a two-stage pipeline for AI-assisted improvement of published algorithm implementations. In the first stage, a large language model with re

model-releasesarxiv-cs-ai
17 Apr 2026
Agents

Timescale Separation Enables Deep Reinforcement Learning Control of Rotating Detonation Engine Mode Transitions

DGX agent

arXiv:2604.14398v1 Announce Type: cross Abstract: Rotating detonation engines (RDEs) are a promising propulsion concept that may offer higher thermodynamic efficiency and specific impulse than convent

agentsarxiv-cs-lg
17 Apr 2026
Agents

ToolSpec: Accelerating Tool Calling via Schema-Aware and Retrieval-Augmented Speculative Decoding

DGX agent

arXiv:2604.13519v1 Announce Type: new Abstract: Tool calling has greatly expanded the practical utility of large language models (LLMs) by enabling them to interact with external applications. As LLM

agentsarxiv-cs-cl
16 Apr 2026
Safety

Dynamic Multi-Robot Task Allocation under Uncertainty and Communication Constraints: A Game-Theoretic Approach

DGX agent

arXiv:2604.11954v1 Announce Type: cross Abstract: We study dynamic multi-robot task allocation under uncertain task completion, time-window constraints, and incomplete information. Tasks arrive online

safetyarxiv-cs-ro
15 Apr 2026
Agents

EMBER: Autonomous Cognitive Behaviour from Learned Spiking Neural Network Dynamics in a Hybrid LLM Architecture

DGX agent

arXiv:2604.12167v1 Announce Type: new Abstract: We present (Experience-Modulated Biologically-inspired Emergent Reasoning), a hybrid cognitive architecture that reorganises the relationship between la

agentsarxiv-cs-ai
15 Apr 2026
Model Releases

Nemotron 3 Super: Open, Efficient Mixture-of-Experts Hybrid Mamba-Transformer Model for Agentic Reasoning

DGX agent

arXiv:2604.12374v1 Announce Type: cross Abstract: We describe the pre-training, post-training, and quantization of Nemotron 3 Super, a 120 billion (active 12 billion) parameter hybrid Mamba-Attention

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ATANT v1.1: Positioning Continuity Evaluation Against Memory, Long-Context, and Agentic-Memory Benchmarks

DGX agent

arXiv:2604.10981v1 Announce Type: new Abstract: ATANT v1.0 (arXiv:2604.06710) defined continuity as a system property with 7 required properties and introduced a 10-checkpoint, LLM-free evaluation met

model-releasesarxiv-cs-ai
14 Apr 2026
Research

Back to Basics: Let Conversational Agents Remember with Just Retrieval and Generation

DGX agent

arXiv:2604.11628v1 Announce Type: new Abstract: Existing conversational memory systems rely on complex hierarchical summarization or reinforcement learning to manage long-term dialogue history, yet re

researcharxiv-cs-cl
14 Apr 2026
Safety

Interactive Learning for LLM Reasoning

DGX agent

arXiv:2509.26306v4 Announce Type: replace Abstract: Existing multi-agent learning approaches have developed interactive training environments to explicitly promote collaboration among multiple Large L

safetyarxiv-cs-ai
14 Apr 2026
Agents

Machine Learning-Based Detection of MCP Attacks

DGX agent

arXiv:2604.10534v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) is a new and emerging technology that extends the functionality of large language models, improving workflows but als

agentsarxiv-cs-ai
14 Apr 2026
Agents

Neuro-Symbolic Strong-AI Robots with Closed Knowledge Assumption: Learning and Deductions

DGX agent

arXiv:2604.09567v1 Announce Type: cross Abstract: Knowledge representation formalisms are aimed to represent general conceptual information and are typically used in the construction of the knowledge

agentsarxiv-cs-ai
14 Apr 2026
Agents

ActivityEditor: Learning to Synthesize Physically Valid Human Mobility

DGX agent

arXiv:2604.05529v2 Announce Type: replace Abstract: Human mobility modeling is indispensable for diverse urban applications. However, existing data-driven methods often suffer from data scarcity, limi

agentsarxiv-cs-ai
13 Apr 2026
Agents

Adversarial Sensor Errors for Safe and Robust Wind Turbine Fleet Control

DGX agent

arXiv:2604.08750v1 Announce Type: new Abstract: Plant-level control is an emerging wind energy technology that presents opportunities and challenges. By controlling turbines in a coordinated manner vi

agentsarxiv-cs-lg
13 Apr 2026
Agents

Beyond Relevance: Utility-Centric Retrieval in the LLM Era

DGX agent

arXiv:2604.08920v1 Announce Type: cross Abstract: Information retrieval systems have traditionally optimized for topical relevance-the degree to which retrieved documents match a query. However, relev

agentsarxiv-cs-ai
13 Apr 2026
Agents

Building Better Environments for Autonomous Cyber Defence

DGX agent

arXiv:2604.08805v1 Announce Type: cross Abstract: In November 2025, the authors ran a workshop on the topic of what makes a good reinforcement learning (RL) environment for autonomous cyber defence (A

agentsarxiv-cs-ai
13 Apr 2026
Research

Exploiting Web Search Tools of AI Agents for Data Exfiltration

DGX agent

arXiv:2510.09093v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are now routinely used to autonomously execute complex tasks, from natural language processing to dynamic workflo

researcharxiv-cs-cl
13 Apr 2026
Tutorials

Mind the Gap Between Spatial Reasoning and Acting! Step-by-Step Evaluation of Agents With Spatial-Gym

DGX agent

arXiv:2604.09338v1 Announce Type: new Abstract: Spatial reasoning is central to navigation and robotics, yet measuring model capabilities on these tasks remains difficult. Existing benchmarks evaluate

tutorialsarxiv-cs-ai
13 Apr 2026
Agents

A Soft Robotic Interface for Chick-Robot Affective Interactions

DGX agent

arXiv:2604.08443v1 Announce Type: new Abstract: The potential of Animal-Robot Interaction (ARI) in welfare applications depends on how much an animal perceives a robotic agent as socially relevant, no

agentsarxiv-cs-ro
10 Apr 2026
Agents

AnyImageNav: Any-View Geometry for Precise Last-Meter Image-Goal Navigation

DGX agent

arXiv:2604.05351v3 Announce Type: replace-cross Abstract: Image Goal Navigation (ImageNav) is evaluated by a coarse success criterion, the agent must stop within 1m of the target, which is sufficient

agentsarxiv-cs-cv
10 Apr 2026
← Previous
1…119120121122123…236
Next →