AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

FALSIFYBENCH: Evaluating Inductive Reasoning in LLMs with Rule Discovery Games

DGX agent

arXiv:2606.04751v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents in scientific tasks. Yet whether these systems can effectively engage in for

agentsarxiv-cs-ai
4 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

INTACT: Ego-Guided Typed Sparse Evidence Retrieval for Heterogeneous Collaborative Perception

DGX agent

arXiv:2606.04437v1 Announce Type: new Abstract: Collaborative perception extends the perceptual range of autonomous vehicles by sharing information across agents, but heterogeneous sensors and percept

agentsarxiv-cs-cv
4 Jun 2026
Agents

TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration

DGX agent

arXiv:2606.04743v1 Announce Type: cross Abstract: Agents are widely deployed as assistants over documents, tools, and code. However, they typically act only on explicit user requests, which surface on

agentsarxiv-cs-ai
4 Jun 2026
Agents

Closed-Loop Molecular Design with Calibrated Deference

DGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

agentsarxiv-cs-ai
3 Jun 2026
Agents

Learning to See via Epiretinal Implant Stimulation in silico with Model-Based Deep Reinforcement Learning

DGX agent

arXiv:2606.03118v1 Announce Type: cross Abstract: Objective: Diseases such as age-related macular degeneration and retinitis pigmentosa cause the degradation of the photoreceptor layer. One approach t

agentsarxiv-cs-cv
3 Jun 2026
Model Releases

ClawHub Security Signals: When VirusTotal, Static Analysis, and SkillSpector Disagree

DGX agent

arXiv:2606.01494v1 Announce Type: cross Abstract: Agent skills extend AI agents with reusable instructions, tools, scripts, references, and workflows, establishing a security boundary distinct from bo

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Interaction-Centered Intelligence: Toward Interaction as the Primary Unit of Analysis in Co-Creative AI and Human-AI Systems

DGX agent

arXiv:2606.00807v1 Announce Type: new Abstract: Traditional artificial intelligence has largely conceptualized intelligence as isolated computation occurring within bounded agents. Across classical AI

agentsarxiv-cs-ai
2 Jun 2026
Model Releases

MindGames Arena Generalization Track: In2AI Solution with Delayed Per-Step Reward Attribution

DGX agent

arXiv:2606.00017v1 Announce Type: new Abstract: Training language model agents for multi-agent strategic interaction presents a core difficulty: the quality of any action may depend on future events t

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Auto-Discovery-Bench: Diagnosing Structured State Tracking in Oracle-Guided Discovery

DGX agent

arXiv:2502.15224v2 Announce Type: replace-cross Abstract: Interactive discovery requires agents to maintain and update structured beliefs over many rounds of feedback. Before evaluating agents in nois

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

Provably Convergent Actor-Critic for MARL through Risk-aversion

DGX agent

arXiv:2602.12386v2 Announce Type: replace-cross Abstract: Learning stationary policies in infinite-horizon general-sum Markov games (MGs) remains a fundamental open problem in Multi-Agent Reinforcemen

agentsarxiv-cs-lg
1 Jun 2026
Model Releases

The Surface You Test Is Not the Surface That Breaks

DGX agent

arXiv:2605.30454v1 Announce Type: cross Abstract: Tool-augmented LLM agents are vulnerable to prompt injection: a third party who controls part of the agent's context can plant instructions that the a

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

Envy-Free Allocation of Indivisible Goods via Noisy Queries

DGX agent

arXiv:2602.06361v2 Announce Type: replace-cross Abstract: We introduce a problem of fairly allocating indivisible goods (items) in which the agents' valuations cannot be observed directly, but instead

agentsarxiv-cs-lg
29 May 2026
Model Releases

GUITestScape: Towards Open-set Evaluation on Exploratory GUI Testing

DGX agent

arXiv:2605.29532v1 Announce Type: cross Abstract: Exploratory GUI testing is a particularly demanding setting for MLLM agents: without predefined test scripts, an agent must autonomously navigate an a

model-releasesarxiv-cs-ai
29 May 2026
Agents

Calibrating Conservatism for Scalable Oversight

DGX agent

arXiv:2605.28807v1 Announce Type: new Abstract: Agentic AI systems capable of autonomous planning and extended environmental interaction pose a fundamental control problem: how can humans maintain mea

agentsarxiv-cs-ai
28 May 2026
Model Releases

TRACER: Turn-level Regret Matching with Inner Reinforcement Credit for Cooperative Multi-LLM Reasoning

DGX agent

arXiv:2605.28699v1 Announce Type: new Abstract: Large language models increasingly rely on either reinforcement learning or multi-agent prompting to improve reasoning, yet these two paradigms remain d

model-releasesarxiv-cs-ai
28 May 2026
Hardware

Uni-LaViRA: Language-Vision-Robot Actions Translation for Unified Embodied Navigation

DGX agent

arXiv:2605.27582v1 Announce Type: cross Abstract: Embodied navigation requires an agent to map language and visual observations to a stream of spatial actions that drive a real robot through environme

hardwarearxiv-cs-cv
28 May 2026
Agents

Decoupled Delay Compensation: Enhancing Pre-trained MARL Policies via Learned Dynamics Filtering

DGX agent

arXiv:2605.26286v1 Announce Type: cross Abstract: Real-world multi-agent reinforcement learning (MARL) systems must often operate under stale observations, stochastic communication delays, and intermi

agentsarxiv-cs-ai
27 May 2026
Agents

Vital Trace: Protocol-Constrained Patient-State Reasoning for Longitudinal Clinical Trajectories

DGX agent

arXiv:2602.12833v2 Announce Type: replace-cross Abstract: Longitudinal clinical reasoning over electronic health records requires tracking evolving physiological measurements, laboratory results, and

agentsarxiv-cs-ai
27 May 2026
Agents

Act or Clarify? Modeling Sensitivity to Uncertainty and Cost in Communication

DGX agent

arXiv:2602.02843v3 Announce Type: replace Abstract: When deciding how to act under uncertainty, agents may choose to act to reduce uncertainty or they may act despite that uncertainty. In communicativ

agentsarxiv-cs-cl
26 May 2026
Agents

Attested Tool-Server Admission: A Security Extension to the Model Context Protocol

DGX agent

arXiv:2605.24248v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) standardizes how a large-language-model (LLM) agent and an external tool server exchange messages, but not trust: a h

agentsarxiv-cs-ai
26 May 2026
Agents

Back to Parsimonious Latents: Learning Task-Centric World Models from Visual Foundations

DGX agent

arXiv:2605.25620v1 Announce Type: new Abstract: World models enable agents to predict future dynamics conditioned on actions, making the choice of latent representation central to planning and control

agentsarxiv-cs-ai
26 May 2026
Agents

Evo-Attacker: Memory-Augmented Reinforcement Learning for Long-Horizon Tool Attacks on LLM-MAS

DGX agent

arXiv:2605.25389v1 Announce Type: cross Abstract: While Large Language Model-based Multi-Agent Systems (LLM-MAS) demonstrate remarkable capabilities in solving complex tasks by orchestrating specializ

agentsarxiv-cs-ai
26 May 2026
Agents

Hybrid Deep Searcher: Scalable Parallel and Sequential Search Reasoning

DGX agent

arXiv:2508.19113v3 Announce Type: replace Abstract: Large reasoning models (LRMs) combined with retrieval-augmented generation (RAG) have enabled deep research agents capable of multi-step reasoning w

agentsarxiv-cs-ai
26 May 2026
Model Releases

Partner-Aware Hierarchical Skill Discovery for Robust Human-AI Collaboration

DGX agent

arXiv:2605.24352v1 Announce Type: new Abstract: Multi-agent collaboration, especially in human-AI teaming, requires agents that can adapt to novel partners with diverse and dynamic behaviors. Conventi

model-releasesarxiv-cs-ai
26 May 2026
Safety

Quantum Frog: Emergent Cooperation and Difficulty Scaling in a Quantized-Time Cooperative Game

DGX agent

arXiv:2605.23930v1 Announce Type: new Abstract: We introduce Quantum Frog, a two-player cooperative game built on a novel quantized-time mechanic in which the environment advances only when a player a

safetyarxiv-cs-ai
26 May 2026
Agents

Toward Enactive Artificial Intelligence

DGX agent

arXiv:2605.24238v1 Announce Type: new Abstract: In this paper, we advocate for incorporating enactive approaches to perception and cognition into artificial intelligence (AI). Enactive approaches view

agentsarxiv-cs-ai
26 May 2026
Model Releases

FastKernels: Benchmarking GPU Kernel Generation in Production

DGX agent

arXiv:2605.23215v1 Announce Type: cross Abstract: LLM-based agents for GPU kernel generation are advancing rapidly, yet their progress is fundamentally constrained by the benchmarks they optimize agai

model-releasesarxiv-cs-ai
25 May 2026
Agents

Latent Cache Flow: Model-to-Model Communication Without Text

DGX agent

arXiv:2605.22863v1 Announce Type: new Abstract: LLM agents today communicate via text, which incurs considerable latency and information loss due to the need to autoregressively decode the sharer mode

agentsarxiv-cs-lg
25 May 2026
Agents

Beyond Majority Voting: LLM Aggregation by Leveraging Higher-Order Information

DGX agent

arXiv:2510.01499v2 Announce Type: replace-cross Abstract: With the rapid progress of multi-agent large language model (LLM) reasoning, how to effectively aggregate answers from multiple LLMs has emerg

agentsarxiv-cs-ai
20 May 2026
Agents

When Individually Calibrated Models Become Collectively Miscalibrated

DGX agent

arXiv:2605.18858v1 Announce Type: cross Abstract: Probabilistic prediction systems often aggregate probability estimates from multiple models into a single decision. A common assumption is that if eac

agentsarxiv-cs-ai
20 May 2026
Model Releases

DexHoldem: Playing Texas Hold'em with Dexterous Embodied System

DGX agent

arXiv:2605.18727v1 Announce Type: cross Abstract: Evaluating embodied systems on real dexterous hardware requires more than isolated primitive skills: an agent must perceive a changing tabletop scene,

model-releasesarxiv-cs-ai
19 May 2026
Safety

Experiment-as-Code Labs: A Declarative Stack for AI-Driven Scientific Discovery

DGX agent

arXiv:2605.04375v2 Announce Type: replace-cross Abstract: To unleash the full potential of AI for Science, we must untether the agents from a purely digital environment. The agent's ability to control

safetyarxiv-cs-ai
19 May 2026
Agents

LASAR: Towards Spatio-temporal Reasoning with Latent Cognitive Map

DGX agent

arXiv:2605.16899v1 Announce Type: new Abstract: A fundamental challenge in embodied AI is verifying if agents build internal models of spatial structure or merely learn to mimic task-specific expert t

agentsarxiv-cs-cv
19 May 2026
Agents

Qumus: Realization of An Embodied AI Quantum Material Experimentalist

DGX agent

arXiv:2605.18407v1 Announce Type: cross Abstract: While modern Large Language Models (LLMs) and agentic artificial intelligence (AI) have demonstrated transformative capabilities in digital domains, t

agentsarxiv-cs-ai
19 May 2026
Safety

DR Tulu: Reinforcement Learning with Evolving Rubrics for Deep Research

DGX agent

arXiv:2511.19399v3 Announce Type: replace-cross Abstract: Deep research agents perform multi-step research to produce long-form, well-attributed answers. However, most open deep research agents are tr

safetyarxiv-cs-ai
18 May 2026
Agents

Shaping Sparse Rewards in Reinforcement Learning: A Semi-supervised Approach

DGX agent

arXiv:2501.19128v5 Announce Type: replace-cross Abstract: In many real-world scenarios, reward signal for agents are exceedingly sparse, making it challenging to learn an effective reward function for

agentsarxiv-cs-ai
18 May 2026
Agents

WorldVLN: Autoregressive World Action Model for Aerial Vision-Language Navigation

DGX agent

arXiv:2605.15964v1 Announce Type: cross Abstract: Aerial vision-language navigation (VLN) requires agents to follow natural-language instructions through closed-loop perception and action in 3D enviro

agentsarxiv-cs-cv
18 May 2026
Agents

X-SYNTH: Beyond Retrieval -- Enterprise Context Synthesis from Observed Human Attention

DGX agent

arXiv:2605.15505v1 Announce Type: new Abstract: In enterprise operations, the context required for an AI agent task is scattered across systems of record, static information stores, and communication

agentsarxiv-cs-ai
18 May 2026
Agents

Closing the Gap on the Sample Complexity of 1-Identification

DGX agent

arXiv:2601.15620v2 Announce Type: replace Abstract: The 1-identification problem is a fundamental pure-exploration problem in multi-armed bandits. An agent aims to determine whether there exists an ar

agentsarxiv-cs-lg
15 May 2026
Agents

Heuristic Pathologies and Further Variance Reduction via Uncertainty Propagation in the AIVAT Family of Techniques

DGX agent

arXiv:2605.14261v1 Announce Type: new Abstract: How should an agent's performance in a multiagent environment be evaluated when there is a limited sample size or a high cost of running a trial? The AI

agentsarxiv-cs-ai
15 May 2026
Agents

In-IDE Toolkit for Developers of AI-Based Features

DGX agent

arXiv:2605.14612v1 Announce Type: cross Abstract: AI-enabled features built on LLMs and agentic workflows are difficult to test, debug, and reproduce, especially for product-focused software engineers

agentsarxiv-cs-ai
15 May 2026
Safety

ERPPO: Entropy Regularization-based Proximal Policy Optimization

DGX agent

arXiv:2605.13131v1 Announce Type: new Abstract: Multi-Agent Proximal Policy Optimization (MAPPO) is a variant of the Proximal Policy Optimization (PPO) algorithm, specifically tailored for multi-agent

safetyarxiv-cs-lg
14 May 2026
Agents

Learning POMDP World Models from Observations with Language-Model Priors

DGX agent

arXiv:2605.13740v1 Announce Type: new Abstract: Whether navigating a building, operating a robot, or playing a game, an agent that acts effectively in an environment must first learn an internal model

agentsarxiv-cs-lg
14 May 2026
Agents

What Limits Vision-and-Language Navigation ?

DGX agent

arXiv:2605.13328v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) is a cornerstone of embodied intelligence. However, current agents often suffer from significant performance degr

agentsarxiv-cs-ai
14 May 2026
Model Releases

PRISM: : Planning and Reasoning with Intent in Simulated Embodied Environments

DGX agent

arXiv:2605.11534v1 Announce Type: new Abstract: When an LLM-based embodied agent fails at a household task, the culprit could be misidentified objects, forgotten sub-goals, or poor action sequencing -

model-releasesarxiv-cs-ro
13 May 2026
Safety

AIPO: : Learning to Reason from Active Interaction

DGX agent

arXiv:2605.08401v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have demonstrated remarkable reasoning capabilities, largely stimulated by Reinforcement Learning with

safetyarxiv-cs-ai
12 May 2026
Agents

HULK: Large-scale Hierarchical Coordination under Continual and Uncertain Temporal Tasks

DGX agent

arXiv:2605.08722v1 Announce Type: new Abstract: Multi-agent systems can be extremely efficient when working concurrently and collaboratively, e.g., for delivery, surveillance, search and rescue. Coord

agentsarxiv-cs-ro
12 May 2026
Agents

MapNav: A Novel Memory Representation via Annotated Semantic Maps for Vision-and-Language Navigation

DGX agent

arXiv:2502.13451v5 Announce Type: replace Abstract: Vision-and-language navigation (VLN) is a key task in Embodied AI, requiring agents to navigate diverse and unseen environments while following natu

agentsarxiv-cs-ro
12 May 2026
← Previous
1…125126127128129…236
Next →