AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

Learning Ordinal Response Policies in Rank-Based Stochastic Prize-Collecting Games

DGX agent

arXiv:2510.24515v2 Announce Type: replace Abstract: The Team Orienteering Problem (TOP) generalizes many real-world multi-agent scheduling and routing tasks that occur in autonomous mobility, aerial l

model-releasesarxiv-cs-ro
11 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

(Auto)formalization is supposed to be easy: Trellis process semantics for spelling out rigorous proofs

DGX agent

arXiv:2606.09674v1 Announce Type: new Abstract: We present Trellis: an autoformalization system that leverages LLM agents in a deterministically constrained workflow to enforce incremental progress in

agentsarxiv-cs-ai
9 Jun 2026
Agents

FASE: Fast Adaptive Semantic Entropy for Code Quality

DGX agent

arXiv:2606.09800v1 Announce Type: cross Abstract: Multi-agent code generation offers a promising paradigm for autonomous software development by simulating the human software engineering lifecycle. Ho

agentsarxiv-cs-ai
9 Jun 2026
Agents

RepoLaunch: Automating Build and Management of Code Repositories across Languages and Platforms

DGX agent

arXiv:2603.05026v2 Announce Type: replace-cross Abstract: Language model (LM) agents have driven substantial progress in automated software engineering (SWE), yet building and testing software reposit

agentsarxiv-cs-lg
9 Jun 2026
Agents

To Nuke or Not to Nuke: LLMs' (Missing) Ethical Reasoning and Actions in a High-Stakes Decision-Making Simulation

DGX agent

arXiv:2606.08310v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as long-horizon agents with decision-making capacities. While LLMs can show ethical competence on

agentsarxiv-cs-ai
9 Jun 2026
Agents

IRAF: Interference-Resilient Adaptive Fusion for Noise-Robust End-to-End Full-Duplex Spoken Dialogue Systems

DGX agent

arXiv:2606.06559v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models allow voice agents to listen and speak concurrently, enabling natural interaction with real-time overlap. However,

agentsarxiv-cs-ai
8 Jun 2026
Agents

A Finite Certificate for the Positive n=9 Vasc Inequality

DGX agent

arXiv:2606.06136v1 Announce Type: cross Abstract: We prove the positive-real n=9 case of the Vasc cyclic inequality. The proof was obtained with human-guided assistance from the AI agent MechMath Agen

agentsarxiv-cs-ai
6 Jun 2026
Agents

The Self-Correction Illusion: LLMs Correct Others but Not Themselves

DGX agent

arXiv:2606.05976v1 Announce Type: cross Abstract: Recent work shows that LLM agents struggle to correct errors in their own reasoning traces yet show markedly higher correction rates when identical cl

agentsarxiv-cs-cl
5 Jun 2026
Agents

Beyond Correctness: Rewarding Faithful Reasoning in Retrieval-Augmented Generation

DGX agent

arXiv:2510.13272v3 Announce Type: replace Abstract: Inspired by the success of reinforcement learning (RL) in Large Language Model (LLM) training for domains like math and code, recent work has begun

agentsarxiv-cs-cl
4 Jun 2026
Agents

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers

DGX agent

arXiv:2606.04421v1 Announce Type: new Abstract: Many current agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure: when an outcome d

agentsarxiv-cs-ai
4 Jun 2026
Agents

CARVE: Certified Affordable Repair of Vetoed Maneuvers via Envelopes for Interactive Driving

DGX agent

arXiv:2606.02641v1 Announce Type: cross Abstract: Interactive driving exposes a failure mode that is easy to miss in rule-aware autonomous-driving stacks: a hard-rule margin can be negative for an ego

agentsarxiv-cs-ai
3 Jun 2026
Agents

From Control Boundary to Insurance Claim: Reconstructing AI-Mediated Losses Through the CER Framework

DGX agent

arXiv:2606.03777v1 Announce Type: new Abstract: AI losses that arise through an insured organization's generative or agentic AI system require state reconstruction, not merely event reconstruction, be

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

KForge: LLM-Driven Cross-Platform Kernel Generation for AI Accelerators

DGX agent

arXiv:2606.02963v1 Announce Type: new Abstract: Production inference increasingly targets a heterogeneous mix of accelerators. Agentic pipelines interleave reasoning, tool calls, and multi-agent coord

model-releasesarxiv-cs-lg
3 Jun 2026
Agents

MemTrain: Self-Supervised Context Memory Training

DGX agent

arXiv:2606.03197v1 Announce Type: new Abstract: Memory is an indispensable capability for long-horizon LLM agents, enabling them to preserve and utilize information accumulated across extended interac

agentsarxiv-cs-cl
3 Jun 2026
Agents

MindClaw: Closed-Loop Embodied Mental-State Reasoning for Precision Intervention

DGX agent

arXiv:2606.01063v1 Announce Type: new Abstract: Theory of Mind (ToM) enables an agent to reason about another actor's beliefs, goals, and intentions, which is essential for human-centered embodied ass

agentsarxiv-cs-ai
2 Jun 2026
Agents

OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence

DGX agent

arXiv:2603.14771v3 Announce Type: replace Abstract: Large Language Model (LLM)-based Collective Intelligence (CI) presents a promising approach to overcoming the data wall and continuously boosting th

agentsarxiv-cs-ai
2 Jun 2026
Agents

SWE-rebench V2: Language-Agnostic SWE Task Collection at Scale

DGX agent

arXiv:2602.23866v2 Announce Type: replace-cross Abstract: Software engineering agents (SWE) are improving rapidly, with recent gains largely driven by reinforcement learning (RL). However, RL training

agentsarxiv-cs-cl
2 Jun 2026
Agents

Don't Fool Me Twice: Adapting to Adversity in the Wild with Experience-Driven Reasoning

DGX agent

arXiv:2605.31119v1 Announce Type: cross Abstract: In robotics, dangers and adversity modes are often embodiment-specific and relative to each agent. A frontier of autonomous mobile robotics is to enab

agentsarxiv-cs-lg
1 Jun 2026
Model Releases

PInVerify: An Offline Embodied Benchmark for Active Instance Verification

DGX agent

arXiv:2605.30639v1 Announce Type: cross Abstract: Embodied agents have made strong progress in navigating to target objects, but reaching the goal vicinity does not guarantee that the agent has found

model-releasesarxiv-cs-ai
1 Jun 2026
Agents

Bosses, Kings, and the Commons: Cooperation Under Power Asymmetry in LLM Societies

DGX agent

arXiv:2605.29062v1 Announce Type: new Abstract: Communities can sustainably manage shared resources (commons) through self-governance and cooperative norms, a central finding of Ostrom's theory of sel

agentsarxiv-cs-cl
29 May 2026
Agents

FLIP: Real-Time and Resilient Formation Planning for Large-Scale DIstributed Swarms via Point Cloud Registration

DGX agent

arXiv:2605.29704v1 Announce Type: new Abstract: Traditional large-scale formation planning either oversimplify the formation representation which leads to poor performance, or they employ complete col

agentsarxiv-cs-ro
29 May 2026
Agents

Formalizing Mathematics at Scale

DGX agent

arXiv:2605.29955v1 Announce Type: new Abstract: We present AutoformBot, a multi-agent system for building an Autoformalized Textbook Library At Scale (Atlas) in Lean 4. AutoformBot orchestrates thousa

agentsarxiv-cs-ai
29 May 2026
Agents

FRUC: Feedforward Dynamic Scene Reconstruction from Uncalibrated Collaborative Driving Views

DGX agent

arXiv:2605.29997v1 Announce Type: new Abstract: We present FRUC, a feed-forward 3D Gaussian splatting framework for dynamic scene reconstruction from uncalibrated collaborative driving views. Existing

agentsarxiv-cs-cv
29 May 2026
Model Releases

Indexing the Unreadable: LLM-Native Recursive Construction and Search of Service Taxonomies

DGX agent

arXiv:2605.29270v1 Announce Type: new Abstract: The era of the Internet of Agents (IoA) is taking shape: LLM agents are expected to fulfill user goals by orchestrating fast-growing populations of Mode

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

LogDx-CI: Benchmarking Log Reduction Tools for LLM Root-Cause Diagnosis

DGX agent

arXiv:2605.28876v1 Announce Type: cross Abstract: CI failure logs are large (median 5k lines, max 200k in this corpus) and noisy. Coding agents that try to debug them depend on an upstream tool to red

model-releasesarxiv-cs-ai
29 May 2026
Model Releases

When Should a Robot Think? Resource-Aware Reasoning via Reinforcement Learning for Embodied Robotic Decision-Making

DGX agent

arXiv:2603.16673v4 Announce Type: replace-cross Abstract: Embodied robotic systems increasingly rely on large language model (LLM)-based agents to support high-level reasoning, planning, and decision-

model-releasesarxiv-cs-ai
29 May 2026
Agents

Reasoning and Planning with Dynamically Changing Norms

DGX agent

arXiv:2605.27622v1 Announce Type: new Abstract: To safely interact with humans, AI agents must both know our norms and consider them during planning. However, such norm-guided planning has been less e

agentsarxiv-cs-ai
28 May 2026
Agents

Rethinking Memory as Continuously Evolving Connectivity

DGX agent

arXiv:2605.28773v1 Announce Type: cross Abstract: Existing memory-augmented LLM agents often treat memory as a static repository with pre-defined representations and fixed retrieval pipelines, which i

agentsarxiv-cs-ai
28 May 2026
Agents

3D Gaussian Map with Open-Set Semantic Grouping for Vision-Language Navigation

DGX agent

arXiv:2605.26500v1 Announce Type: new Abstract: Vision-language navigation (VLN) requires an agent to traverse complex 3D environments based on natural language instructions, necessitating a thorough

agentsarxiv-cs-cv
27 May 2026
Model Releases

Sentinel: Embodied Cooperative Spatial Reasoning and Planning

DGX agent

arXiv:2605.26239v1 Announce Type: new Abstract: In this work, we study Cooperative Spatial Intelligence, the ability of decentralized embodied agents to coordinate effectively under dynamic environmen

model-releasesarxiv-cs-cv
27 May 2026
Hardware

SIA: Self Improving AI with Harness & Weight Updates

DGX agent

arXiv:2605.27276v1 Announce Type: new Abstract: Humans are the bottleneck in building and improving AI. Both the models and the agents that wrap them are written, tuned, and corrected by people. The l

hardwarearxiv-cs-ai
27 May 2026
Tutorials

A perspective on fluid mechanical environments for challenges in reinforcement learning

DGX agent

arXiv:2605.25011v1 Announce Type: new Abstract: We consider the challenge of developing agents that efficiently interact with high-dimensional, evolving environments, towards a view of practical reinf

tutorialsarxiv-cs-lg
26 May 2026
Model Releases

CausaLab: A Scalable Environment for Interactive Causal Discovery Toward AI Scientists

DGX agent

arXiv:2605.26029v1 Announce Type: new Abstract: We introduce CausaLab, a scalable environment for evaluating interactive causal discovery by LLM agents. Unlike prior evaluations, CausaLab evaluates bo

model-releasesarxiv-cs-ai
26 May 2026
Agents

MCPXKIT: The Unified Toolkit for Analyzing Model Context Protocol Security

DGX agent

arXiv:2508.12538v2 Announce Type: replace-cross Abstract: The Model Context Protocol (MCP) has emerged as a universal standard that enables AI agents to seamlessly connect with external tools, signifi

agentsarxiv-cs-ai
26 May 2026
Agents

Reward Shaping and Action Masking for Compositional Tasks using Behavior Trees and LLMs

DGX agent

arXiv:2605.05795v2 Announce Type: replace Abstract: Decomposing complex tasks into a sequence of simpler subtasks can improve learning efficiency for an autonomous agent. Reinforcement learning (RL) c

agentsarxiv-cs-lg
26 May 2026
Model Releases

SMDD-Bench: Can LLMs Solve Real-World Small Molecule Drug Design Tasks?

DGX agent

arXiv:2605.21740v2 Announce Type: replace Abstract: LLM agents have incredible potential for scientific discovery applications. However, the performance of LLM agents on real-world, small molecule dru

model-releasesarxiv-cs-ai
26 May 2026
Agents

Socially fluent AI decouples conversational signals from source identity in online interaction

DGX agent

arXiv:2605.23426v1 Announce Type: cross Abstract: Socially fluent agentic AI can now participate in online interaction in ways that resemble ordinary human conversation, potentially weakening people's

agentsarxiv-cs-ai
25 May 2026
Agents

AwareVLN: Reasoning with Self-awareness for Vision-Language Navigation

DGX agent

arXiv:2605.22816v1 Announce Type: cross Abstract: Vision-and-Language Navigation (VLN) requires an agent to ground language instructions to its own movement within a visual environment. While state-of

agentsarxiv-cs-cv
22 May 2026
Safety

Diagnosis Is Not Prescription: Linguistic Co-Adaptation Explains Patching Hazards in LLM Pipelines

DGX agent

arXiv:2605.21958v1 Announce Type: new Abstract: When a multi-module LLM agent fails, the module most responsible for the failure is not necessarily the best place to intervene. We demonstrate this Dia

safetyarxiv-cs-cl
22 May 2026
Agents

ProCrit: Self-Elicited Multi-Perspective Reasoning with Critic-Guided Revision for Multimodal Sarcasm Detection

DGX agent

arXiv:2605.20867v1 Announce Type: cross Abstract: Multimodal sarcasm detection requires reasoning over cross-modal incongruities between literal expression and intended meaning, yet the specific analy

agentsarxiv-cs-cv
21 May 2026
Model Releases

How Far Are We From True Auto-Research?

DGX agent

arXiv:2605.19156v1 Announce Type: new Abstract: Recent auto-research systems can produce complete papers, but feasibility is not the same as quality, and the field still lacks a systematic study of ho

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

A Machine with Short-Term, Episodic, and Semantic Memory Systems

DGX agent

arXiv:2212.02098v5 Announce Type: replace Abstract: Inspired by the cognitive science theory of the explicit human memory systems, we have modeled an agent with short-term, episodic, and semantic memo

model-releasesarxiv-cs-ai
19 May 2026
Agents

Convergence of Multiagent Learning Systems for Traffic control

DGX agent

arXiv:2511.11654v2 Announce Type: replace-cross Abstract: Rapid urbanization in cities like Bangalore has led to severe traffic congestion, making efficient Traffic Signal Control (TSC) essential. Mul

agentsarxiv-cs-ai
19 May 2026
Agents

Learning to Learn from Multimodal Experience

DGX agent

arXiv:2605.16857v1 Announce Type: new Abstract: Experience-driven learning has emerged as a promising paradigm for enabling agents to improve from interaction trajectories by accumulating and reusing

agentsarxiv-cs-ai
19 May 2026
Agents

PersonaArena: Dynamic Simulation for Evaluating and Enhancing Persona-Level Role-Playing in Large Language Models

DGX agent

arXiv:2605.17044v1 Announce Type: new Abstract: Large language models (LLMs) increasingly serve as interactive social agents, yet their ability to maintain coherent and authentic persona-level role-pl

agentsarxiv-cs-ai
19 May 2026
Agents

Task-Level AI Readiness Assessment for Business Process Management:The T-IPO Model and LARA Matrix in Financial-Services IT Operations

DGX agent

arXiv:2605.16297v1 Announce Type: cross Abstract: Which tasks inside an enterprise workflow can a large-language-model agent reliably handle, and under what conditions? Most business process modeling

agentsarxiv-cs-ai
19 May 2026
Agents

ICRL: Learning to Internalize Self-Critique with Reinforcement Learning

DGX agent

arXiv:2605.15224v1 Announce Type: new Abstract: Large language model-based agents make mistakes, yet critique can often guide the same model toward correct behavior. However, when critique is removed,

agentsarxiv-cs-ai
18 May 2026
Agents

Mecha-nudges for Machines

DGX agent

arXiv:2603.23433v2 Announce Type: replace Abstract: AI agents are becoming active decision-makers on the Internet. As they make decisions in the same environments as humans, the environments themselve

agentsarxiv-cs-ai
18 May 2026
← Previous
1…121122123124125…236
Next →