AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Safety

Perception with Guarantees: Certified Pose Estimation via Reachability Analysis

DGX agent

arXiv:2602.10032v2 Announce Type: replace Abstract: Agents in cyber-physical systems are increasingly entrusted with safety-critical tasks. Ensuring safety of these agents often requires localizing th

safetyarxiv-cs-cv
14 May 2026
Agents

State-Centric Decision Process

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2605.12755v1 Announce Type: new Abstract: Language environments such as web browsers, code terminals, and interactive simulations emit raw text rather than states, and provide none of the runtim

agentsarxiv-cs-ai
14 May 2026
Model Releases

Do Enterprise Systems Need Learned World Models? The Importance of Context to Infer Dynamics

DGX agent

arXiv:2605.12178v1 Announce Type: cross Abstract: World models enable agents to anticipate the effects of their actions by internalizing environment dynamics. In enterprise systems, however, these dyn

model-releasesarxiv-cs-cl
13 May 2026
Agents

Stochastic Minimum-Cost Reach-Avoid Reinforcement Learning

DGX agent

arXiv:2605.11975v1 Announce Type: new Abstract: We study stochastic minimum-cost reach-avoid reinforcement learning, where an agent must satisfy a reach-avoid specification with probability at least p

agentsarxiv-cs-lg
13 May 2026
Agents

DataMaster: Towards Autonomous Data Engineering for Machine Learning

DGX agent

arXiv:2605.10906v1 Announce Type: cross Abstract: As model families, training recipes, and compute budgets become increasingly standardized, further gains in machine learning systems depend increasing

agentsarxiv-cs-ai
12 May 2026
Safety

Emergence of Physical Intelligence via Controllable Information Production

DGX agent

arXiv:2601.22449v2 Announce Type: replace Abstract: Intrinsic Motivation (IM) aims to train agents without external rewards, enabling useful behavior to emerge from the agent's interaction with its en

safetyarxiv-cs-ai
12 May 2026
Agents

Evidence Over Plans: Online Trajectory Verification for Skill Distillation

DGX agent

arXiv:2605.09192v1 Announce Type: new Abstract: Agent skills can remarkably improve task success rates by using human-written procedural documents, but their quality is difficult to assess without env

agentsarxiv-cs-ai
12 May 2026
Agents

Optimal and Scalable MAPF via Multi-Marginal Optimal Transport and Schrodinger Bridges

DGX agent

arXiv:2605.10917v1 Announce Type: new Abstract: We consider anonymous multi-agent path finding (MAPF) where a set of robots is tasked to travel to a set of targets on a finite, connected graph. We sho

agentsarxiv-cs-lg
12 May 2026
Agents

PRISM: Fast Online LLM Serving via Scheduling-Memory Co-design

DGX agent

arXiv:2605.08581v1 Announce Type: new Abstract: Modern online large language model (LLM) services, such as Retrieval-Augmented Generation (RAG) and agent systems, increasingly expose two prominent cha

agentsarxiv-cs-lg
12 May 2026
Model Releases

Process Matters more than Output for Distinguishing Humans from Machines

DGX agent

arXiv:2605.06524v2 Announce Type: replace Abstract: Reliable human-machine discrimination is becoming increasingly important as large language models and autonomous agents are deployed in online setti

model-releasesarxiv-cs-ai
12 May 2026
Hardware

SkillEvolver: Skill Learning as a Meta-Skill

DGX agent

arXiv:2605.10500v1 Announce Type: new Abstract: Agent skills today are static artifact: authored once -- by human curation or one-shot generation from parametric knowledge -- and then consumed unchang

hardwarearxiv-cs-ai
12 May 2026
Agents

The First Drop of Ink: Nonlinear Impact of Misleading Information in Long-Context Reasoning

DGX agent

arXiv:2605.10828v1 Announce Type: new Abstract: As large language models are increasingly deployed in retrieval-augmented generation and agentic systems that accumulate extensive context, understandin

agentsarxiv-cs-ai
12 May 2026
Agents

The Reciprocity Gradient

DGX agent

arXiv:2605.08323v1 Announce Type: cross Abstract: Communication is fundamental to sustaining reciprocity and cooperation in strategic interactions. We identify and formulate the influence attribution

agentsarxiv-cs-ai
12 May 2026
Agents

TinyTroupe: An LLM-powered Multiagent Persona Simulation Toolkit

DGX agent

arXiv:2507.09788v2 Announce Type: replace-cross Abstract: Recent advances in Large Language Models (LLM) have led to a new class of autonomous agents, renewing and expanding interest in the area. LLM-

agentsarxiv-cs-ai
12 May 2026
Agents

CASCADE: Case-Based Continual Adaptation for Large Language Models During Deployment

DGX agent

arXiv:2605.06702v1 Announce Type: new Abstract: Large language models (LLMs) have become a central foundation of modern artificial intelligence, yet their lifecycle remains constrained by a rigid sepa

agentsarxiv-cs-ai
11 May 2026
Agents

Can LLMs Make (Personalized) Access Control Decisions?

DGX agent

arXiv:2511.20284v2 Announce Type: replace-cross Abstract: Precise access control decisions are crucial for the security of both traditional applications and emerging agent-based systems. Typically, th

agentsarxiv-cs-ai
7 May 2026
Safety

Overcoming Environmental Meta-Stationarity in MARL via Adaptive Curriculum and Counterfactual Group Advantage

DGX agent

arXiv:2506.07548v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) has reached competitive performance on cooperative tasks against scripted adversaries, yet most meth

safetyarxiv-cs-ro
7 May 2026
Agents

APIOT: Autonomous Vulnerability Management Across Bare-Metal Industrial OT Networks

DGX agent

arXiv:2605.02346v1 Announce Type: cross Abstract: Bare-metal operational technology (OT) devices -- especially the microcontrollers running Modbus/TCP and CoAP at the base of industrial control system

agentsarxiv-cs-ai
6 May 2026
Agents

Aura-CAPTCHA: A Reinforcement Learning and GAN-Enhanced Multi-Modal CAPTCHA System

DGX agent

arXiv:2508.14976v2 Announce Type: replace Abstract: We present Aura-CAPTCHA, a multi-modal verification system that integrates Generative Adversarial Networks (GANs), Reinforcement Learning (RL), and

agentsarxiv-cs-lg
6 May 2026
Research

Evaluating Generative Models as Interactive Emergent Representations of Human-Like Collaborative Behavior

DGX agent

arXiv:2605.03855v1 Announce Type: new Abstract: Human-AI collaboration requires AI agents to understand human behavior for effective coordination. While advances in foundation models show promising ca

researcharxiv-cs-ro
6 May 2026
Agents

GhostServe: A Lightweight Checkpointing System in the Shadow for Fault-Tolerant LLM Serving

DGX agent

arXiv:2605.00831v1 Announce Type: cross Abstract: The rise of million-token, agent-based applications has placed unprecedented demands on large language model (LLM) inference services. The long-runnin

agentsarxiv-cs-ai
6 May 2026
Agents

HepScript: A Dual-Use DSL for Human-AI Collaborative Data Analysis Workflows in High-Energy Physics

DGX agent

arXiv:2605.01423v1 Announce Type: cross Abstract: The escalating data scale in High-Energy Physics (HEP) fuels a growing aspiration for higher analytical efficiency. While Large Language Models (LLMs)

agentsarxiv-cs-ai
6 May 2026
Agents

Latent State Design for World Models under Sufficiency Constraints

DGX agent

arXiv:2605.01694v1 Announce Type: new Abstract: A world model matters to an agent only through the state it constructs. That state must preserve some information, discard other information, and suppor

agentsarxiv-cs-ai
6 May 2026
Agents

Neural Control: Adjoint Learning Through Equilibrium Constraints

DGX agent

arXiv:2605.03288v1 Announce Type: new Abstract: Many physical AI tasks are governed by implicit equilibrium: an agent actuates a subset of degrees of freedom (boundary DoFs), while the remaining free

agentsarxiv-cs-ro
6 May 2026
Agents

AutoFocus: Uncertainty-Aware Active Visual Search for GUI Grounding

DGX agent

arXiv:2605.02630v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have enabled autonomous GUI agents that translate natural language instructions into executable screen coordinates. Howeve

agentsarxiv-cs-cv
5 May 2026
Agents

Ergodic Risk Measures: Towards a Risk-Aware Foundation for Continual Reinforcement Learning

DGX agent

arXiv:2510.02945v3 Announce Type: replace Abstract: Continual reinforcement learning (continual RL) seeks to formalize the notions of lifelong learning and endless adaptation in RL. In particular, the

agentsarxiv-cs-lg
5 May 2026
Agents

S^3-R1: Learning to Retrieve and Answer Step-by-Step with Synthetic Data

DGX agent

arXiv:2605.01248v1 Announce Type: new Abstract: Reinforcement learning (RL) post-training has enabled newer capabilities in models, such as agentic tool-use for search. However, these models struggle

agentsarxiv-cs-lg
5 May 2026
Agents

The Reasoning Trap: An Information-Theoretic Bound on Closed-System Multi-Step LLM Reasoning

DGX agent

arXiv:2605.01704v1 Announce Type: new Abstract: When copies of the same language model are prompted to debate, they produce diverse phrasings of one perspective rather than diverse perspectives. Multi

agentsarxiv-cs-cl
5 May 2026
Agents

To Call or Not to Call: A Framework to Assess and Optimize LLM Tool Calling

DGX agent

arXiv:2605.00737v1 Announce Type: new Abstract: Agentic AI architectures augment LLMs with external tools, unlocking strong capabilities. However, tool use is not always beneficial; some calls may be

agentsarxiv-cs-ai
5 May 2026
Agents

Pessimism-Free Offline Learning in General-Sum Games via KL Regularization

DGX agent

arXiv:2605.00264v1 Announce Type: new Abstract: Offline multi-agent reinforcement learning in general-sum settings is challenged by the distribution shift between logged datasets and target equilibriu

agentsarxiv-cs-lg
4 May 2026
Safety

PALCAS: A Priority-Aware Intelligent Lane Change Advisory System for Autonomous Vehicles using Federated Reinforcement Learning

DGX agent

arXiv:2604.27118v1 Announce Type: cross Abstract: We present a priority-aware intelligent lane change advisory system based on multi-agent federated reinforcement learning, namely PALCAS, for autonomo

safetyarxiv-cs-ai
1 May 2026
Agents

SpaAct: Spatially-Activated Transition Learning with Curriculum Adaptation for Vision-Language Navigation

DGX agent

arXiv:2604.27620v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) aims to enable an embodied agent to follow natural-language instructions and navigate to a target location in unsee

agentsarxiv-cs-cv
1 May 2026
Model Releases

The TEA Nets framework combines AI and cognitive network science to model targets, events and actors in text

DGX agent

arXiv:2604.27673v1 Announce Type: new Abstract: We introduce Target-Event-Agent Networks (TEA Nets) as a computational framework to extract subjects (``Agents'), verbs (``Events'), and objects (``Targ

model-releasesarxiv-cs-ai
1 May 2026
Safety

Lifting Embodied World Models for Planning and Control

DGX agent

arXiv:2604.26182v1 Announce Type: cross Abstract: World models of embodied agents predict future observations conditioned on an action taken by the agent. For complex embodiments, action spaces are hi

safetyarxiv-cs-ai
30 Apr 2026
Agents

Principled Learning-to-Communicate with Quasi-Classical Information Structures

DGX agent

arXiv:2603.03664v2 Announce Type: replace-cross Abstract: Learning-to-communicate (LTC) in partially observable environments has received increasing attention in deep multi-agent reinforcement learnin

agentsarxiv-cs-lg
30 Apr 2026
Agents

RL unknotter, hard unknots and unknotting number

DGX agent

arXiv:2603.07955v3 Announce Type: replace-cross Abstract: We develop a reinforcement learning pipeline for simplifying knot diagrams. A trained agent learns move proposals and a value heuristic for na

agentsarxiv-cs-lg
30 Apr 2026
Model Releases

AutoGUI-v2: A Comprehensive Multi-Modal GUI Functionality Understanding Benchmark

DGX agent

arXiv:2604.24441v1 Announce Type: new Abstract: Autonomous agents capable of navigating Graphical User Interfaces (GUIs) hold the potential to revolutionize digital productivity. However, achieving tr

model-releasesarxiv-cs-cv
28 Apr 2026
Agents

Behavioral Intelligence Platforms: From Event Streams to Autonomous Insight via Probabilistic Journey Graphs, Behavioral Knowledge Extraction, and Grounded Language Generation

DGX agent

arXiv:2604.22762v1 Announce Type: cross Abstract: Contemporary product analytics systems require users to pose explicit queries, such as writing SQL, configuring dashboards, or constructing funnels, b

agentsarxiv-cs-ai
28 Apr 2026
Agents

PhySE: A Psychological Framework for Real-Time AR-LLM Social Engineering Attacks

DGX agent

arXiv:2604.23148v1 Announce Type: new Abstract: The emerging threat of AR-LLM-based Social Engineering (AR-LLM-SE) attacks (e.g. SEAR) poses a significant risk to real-world social interactions. In su

agentsarxiv-cs-ai
28 Apr 2026
Safety

Algorithmic Feature Highlighting for Human-AI Decision-Making

DGX agent

arXiv:2604.22236v1 Announce Type: cross Abstract: Human decision-makers often face choices about complex cases with many potentially relevant features, but limited bandwidth to inspect and integrate a

safetyarxiv-cs-lg
27 Apr 2026
Agents

An Efficient Real-Time Planning Method for Swarm Robotics Based on an Optimal Virtual Tube

DGX agent

arXiv:2505.01380v2 Announce Type: replace Abstract: Robot swarms navigating through unknown obstacle environments are an emerging research area that faces challenges. Performing tasks in such environm

agentsarxiv-cs-ro
27 Apr 2026
Agents

Logistic Bandits with ilde{O}(sqrt{dT}) Regret without Context Diversity Assumptions

DGX agent

arXiv:2604.22161v1 Announce Type: new Abstract: We study the K-armed logistic bandit problem, where at each round, the agent observes K feature vectors associated with K actions. Existing approaches t

agentsarxiv-cs-lg
27 Apr 2026
Agents

DiagramBank: A Large-scale Dataset of Diagram Design Exemplars with Paper Metadata for Retrieval-Augmented Generation

DGX agent

arXiv:2604.20857v1 Announce Type: cross Abstract: Recent advances in autonomous ``AI scientist'' systems have demonstrated the ability to automatically write scientific manuscripts and codes with exec

agentsarxiv-cs-ai
24 Apr 2026
Safety

HARBOR: Automated Harness Optimization

DGX agent

arXiv:2604.20938v1 Announce Type: cross Abstract: Long-horizon language-model agents are dominated, in lines of code and in operational complexity, not by their underlying model but by the harness tha

safetyarxiv-cs-ai
24 Apr 2026
Model Releases

Using Machine Mental Imagery for Representing Common Ground in Situated Dialogue

DGX agent

arXiv:2604.21144v1 Announce Type: cross Abstract: Situated dialogue requires speakers to maintain a reliable representation of shared context rather than reasoning only over isolated utterances. Curre

model-releasesarxiv-cs-ai
24 Apr 2026
Agents

A Field Guide to Decision Making

DGX agent

arXiv:2604.20669v1 Announce Type: cross Abstract: High-consequence decision making demands peak performance from individuals in positions of responsibility. Such executive authority bears the obligati

agentsarxiv-cs-ai
23 Apr 2026
Model Releases

Coding with Eyes: Visual Feedback Unlocks Reliable GUI Code Generating and Debugging

DGX agent

arXiv:2604.19750v1 Announce Type: cross Abstract: Recent advances in Large Language Model (LLM)-based agents have shown remarkable progress in code generation. However, current agent methods mainly re

model-releasesarxiv-cs-ai
23 Apr 2026
Agents

HaS: Accelerating RAG through Homology-Aware Speculative Retrieval

DGX agent

arXiv:2604.20452v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG) expands the knowledge boundary of large language models (LLMs) at inference by retrieving external documents as c

agentsarxiv-cs-cl
23 Apr 2026
← Previous
1…131132133134135…236
Next →