AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable Citations

DGX agent

arXiv:2605.00846v1 Announce Type: new Abstract: Clinical diagnosis requires answers that are accurate, verifiable, and explicitly grounded in official guidelines. While large language models excel at

agentsarxiv-cs-ai
6 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Complexity Horizons of Compressed Models in Analog Circuit Analysis

DGX agent

arXiv:2605.02285v1 Announce Type: new Abstract: The deployment of Large Language Models (LLMs) for specialized engineering domains, such as circuit analysis, often faces a trade-off between reasoning

agentsarxiv-cs-ai
6 May 2026
Model Releases

MAP-Law: Coverage-Driven Retrieval Control for Multi-Turn Legal Consultation

DGX agent

arXiv:2605.01486v1 Announce Type: new Abstract: Legal consultation is a high-stakes, knowledge-intensive task that requires agents to identify relevant legal issues, retrieve authoritative support, an

model-releasesarxiv-cs-ai
6 May 2026
Agents

S^2tory: Story Spine Distillation for Movie Script Summarization

DGX agent

arXiv:2605.03244v1 Announce Type: new Abstract: Movie scripts pose a fundamental challenge for automatic summarization due to their non-linear, cross-cut narrative structure, which makes surface-level

agentsarxiv-cs-cl
6 May 2026
Agents

SOAR: Real-Time Joint Optimization of Order Allocation and Robot Scheduling in Robotic Mobile Fulfillment Systems

DGX agent

arXiv:2605.03842v1 Announce Type: cross Abstract: Robotic Mobile Fulfillment Systems (RMFS) rely on mobile robots for automated inventory transportation, coordinating order allocation and robot schedu

agentsarxiv-cs-ro
6 May 2026
Agents

WMF-AM: Probing LLM Working Memory via Depth-Parameterized Cumulative State Tracking

DGX agent

arXiv:2603.27343v2 Announce Type: replace Abstract: Existing large language models (LLMs) evaluations use fixed-difficulty benchmarks that cannot adapt as models improve, and rarely isolate specific c

agentsarxiv-cs-ai
6 May 2026
Agents

Adversarial Flow Matching for Imperceptible Attacks on End-to-End Autonomous Driving

DGX agent

arXiv:2605.00880v1 Announce Type: new Abstract: Autonomous driving (AD) is evolving towards end-to-end (E2E) frameworks through two primary paradigms: monolithic models exemplified by Vision-Language-

agentsarxiv-cs-cv
5 May 2026
Agents

AVI-Edit: Audio-sync Video Instance Editing with Granularity-Aware Mask Refiner

DGX agent

arXiv:2512.10571v4 Announce Type: replace Abstract: Recent advancements in video generation highlight that realistic audio-visual synchronization is crucial for engaging content creation. However, exi

agentsarxiv-cs-cv
5 May 2026
Agents

Compiling Deterministic Structure into SLM Harnesses

DGX agent

arXiv:2604.17450v2 Announce Type: replace Abstract: Enterprise SLM deployment faces epistemic asymmetry: small models cannot self-correct reasoning errors, while frontier LLMs incur prohibitive costs

agentsarxiv-cs-ai
5 May 2026
Tutorials

Forager: a lightweight testbed for continual learning with partial observability in RL

DGX agent

arXiv:2605.01131v1 Announce Type: new Abstract: In continual reinforcement learning (CRL), good performance requires never-ending learning, acting, and exploration in a big, partially observable world

tutorialsarxiv-cs-lg
5 May 2026
Agents

Hallucinations Undermine Trust; Metacognition is a Way Forward

DGX agent

arXiv:2605.01428v1 Announce Type: new Abstract: Despite significant strides in factual reliability, errors -- often termed hallucinations -- remain a major concern for generative AI, especially as LLM

agentsarxiv-cs-cl
5 May 2026
Agents

Large Language Models for Multi-Robot Systems: A Survey

DGX agent

arXiv:2502.03814v5 Announce Type: replace Abstract: The rapid advancement of Large Language Models (LLMs) has opened new possibilities in Multi-Robot Systems (MRS), enabling enhanced communication, ta

agentsarxiv-cs-ro
5 May 2026
Agents

LLM Ghostbusters: Surgical Hallucination Suppression via Adaptive Unlearning

DGX agent

arXiv:2605.01047v1 Announce Type: cross Abstract: Hallucinations, outputs that sound plausible but are factually incorrect, remain an open challenge for deployed LLMs. In code generation, models frequ

agentsarxiv-cs-cl
5 May 2026
Agents

Lost in the Tower of Babel: The Adverse Effects of Incidental Multilingualism in LLMs

DGX agent

arXiv:2605.01224v1 Announce Type: new Abstract: This paper argues that contemporary multilingual NLP has converged on a fragile and misleading paradigm of incidental multilingualism. Today's LLMs appe

agentsarxiv-cs-cl
5 May 2026
Agents

Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning

DGX agent

arXiv:2511.03724v3 Announce Type: replace Abstract: AI researchers have long focused on poker-like games as a testbed for environments characterized by multi-player dynamics, imperfect information, an

agentsarxiv-cs-ai
5 May 2026
Safety

Reinforcement Learning from Compiler and Language Server Feedback

DGX agent

arXiv:2510.22907v2 Announce Type: replace Abstract: Coding agents fail when text-level guesses outrun program facts: they hallucinate APIs, drift to the wrong symbol, and apply edits without evidence

safetyarxiv-cs-cl
5 May 2026
Safety

Safe Planning in Interactive Environments via Iterative Policy Updates and Adversarially Robust Conformal Prediction

DGX agent

arXiv:2511.10586v2 Announce Type: replace-cross Abstract: Safe planning of an autonomous agent in interactive environments -- such as the control of a self-driving vehicle among pedestrians -- poses a

safetyarxiv-cs-ro
5 May 2026
Agents

Verbal-R3: Verbal Reranker as the Missing Bridge between Retrieval and Reasoning

DGX agent

arXiv:2605.01399v1 Announce Type: new Abstract: The conventional Retrieval-Augmented Generation (RAG) paradigm of injecting raw retrieved texts into the Large Language Model (LLM)'s context often resu

agentsarxiv-cs-cl
5 May 2026
Safety

BOLT: Online Lightweight Adaptation for Preparation-Free Heterogeneous Cooperative Perception

DGX agent

arXiv:2605.00405v1 Announce Type: new Abstract: Most existing heterogeneous cooperative perception methods depend on prior preparation like offline joint training or tailored collaborator-model adapta

safetyarxiv-cs-cv
4 May 2026
Agents

Causality-enhanced Decision-Making for Autonomous Mobile Robots in Dynamic Environments

DGX agent

arXiv:2504.11901v5 Announce Type: replace Abstract: The growing integration of robots in shared environments-such as warehouses, shopping centres, and hospitals-demands a deep understanding of the und

agentsarxiv-cs-ro
4 May 2026
Model Releases

High-Probability Convergence in Decentralized Stochastic Optimization with Gradient Tracking

DGX agent

arXiv:2605.00281v1 Announce Type: new Abstract: We study high-probability (HP) convergence guarantees in decentralized stochastic optimization, where multiple agents collaborate to jointly train a mod

model-releasesarxiv-cs-lg
4 May 2026
Agents

Long-Horizon Model-Based Offline Reinforcement Learning Without Explicit Conservatism

DGX agent

arXiv:2512.04341v3 Announce Type: replace Abstract: Popular offline reinforcement learning (RL) methods rely on explicit conservatism, penalizing out-of-dataset actions or restricting rollout horizons

agentsarxiv-cs-lg
4 May 2026
Safety

PORTool: Importance-Aware Policy Optimization with Rewarded Tree for Multi-Tool-Integrated Reasoning

DGX agent

arXiv:2510.26020v2 Announce Type: replace Abstract: Multi-tool-integrated reasoning enables LLM-empowered tool-use agents to solve complex tasks by interleaving natural-language reasoning with calls t

safetyarxiv-cs-cl
4 May 2026
Agents

ToolGrad: Efficient Tool-use Dataset Generation with Textual 'Gradients'

DGX agent

arXiv:2508.04086v2 Announce Type: replace Abstract: Prior work synthesizes tool-use LLM datasets by first generating a user query, followed by complex tool-use annotations like depth-first search (DFS

agentsarxiv-cs-cl
4 May 2026
Local Ai

A Collective Variational Principle Unifying Bayesian Inference, Game Theory, and Thermodynamics

DGX agent

arXiv:2604.27942v1 Announce Type: new Abstract: Collective intelligence emerges across biological, physical, and artificial systems without central coordination, yet a unifying principle governing suc

local-aiarxiv-cs-ai
1 May 2026
Agents

Can AI Be a Good Peer Reviewer? A Survey of Peer Review Process, Evaluation, and the Future

DGX agent

arXiv:2604.27924v1 Announce Type: cross Abstract: Peer review is a multi-stage process involving reviews, rebuttals, meta-reviews, final decisions, and subsequent manuscript revisions. Recent advances

agentsarxiv-cs-ai
1 May 2026
Agents

Can AI be a moral victim? The role of moral patiency and ownership perceptions in ethical judgments of using AI-generated content

DGX agent

arXiv:2604.26956v1 Announce Type: cross Abstract: The growing use of generative AI raises ethical concerns about authorship and plagiarism. This study examines how people judge the reuse of AI-generat

agentsarxiv-cs-ai
1 May 2026
Safety

Dreaming Across Towns: Semantic Rollout and Town-Adversarial Regularization for Zero-Shot Held-Out-Town Fixed-Route Driving in CARLA

DGX agent

arXiv:2604.27994v1 Announce Type: new Abstract: Learned driving agents often degrade when deployed in unseen environments. This paper studies a deliberately bounded instance of that problem in the CAR

safetyarxiv-cs-ro
1 May 2026
Agents

Interval Orders, Biorders and Credibility-limited Belief Revision

DGX agent

arXiv:2604.27156v1 Announce Type: new Abstract: Rational belief revision is commonly viewed as being based on a preference order between possible worlds, with the resulting new belief set being those

agentsarxiv-cs-ai
1 May 2026
Agents

Let's Measure Information Step-by-Step: AI-Based Evaluation Beyond Vibes

DGX agent

arXiv:2508.05469v3 Announce Type: replace Abstract: We evaluate artificial intelligence (AI) systems without ground truth by exploiting a link between strategic gaming and information loss. Building o

agentsarxiv-cs-lg
1 May 2026
Agents

OptimusKG: Unifying biomedical knowledge in a modern multimodal graph

DGX agent

arXiv:2604.27269v1 Announce Type: new Abstract: Biomedical knowledge graphs (KGs) are widely used in the life sciences, yet many are derived from unstructured documents and therefore lack schema-level

agentsarxiv-cs-ai
1 May 2026
Safety

TRUST: A Framework for Decentralized AI Service v.0.1

DGX agent

arXiv:2604.27132v1 Announce Type: new Abstract: Large Reasoning Models (LRMs) and Multi-Agent Systems (MAS) in high-stakes domains demand reliable verification, yet centralized approaches suffer four

safetyarxiv-cs-ai
1 May 2026
Agents

3D Generation for Embodied AI and Robotic Simulation: A Survey

DGX agent

arXiv:2604.26509v1 Announce Type: cross Abstract: Embodied AI and robotic systems increasingly depend on scalable, diverse, and physically grounded 3D content for simulation-based training and real-wo

agentsarxiv-cs-cv
30 Apr 2026
Agents

Autonomous Knowledge Graph Exploration with Adaptive Breadth-Depth Retrieval

DGX agent

arXiv:2601.13969v2 Announce Type: replace Abstract: Retrieving evidence for language model queries from knowledge graphs requires balancing broad search across the graph with multi-hop traversal to fo

agentsarxiv-cs-ai
30 Apr 2026
Model Releases

Benchmarks for Trajectory Safety Evaluation and Diagnosis in OpenClaw and Codex: ATBench-Claw and ATBench-Codex

DGX agent

arXiv:2604.14858v2 Announce Type: replace Abstract: As agent systems move into increasingly diverse execution settings, trajectory-level safety evaluation and diagnosis require benchmarks that evolve

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

Lyapunov-Guided Self-Alignment: Test-Time Adaptation for Offline Safe Reinforcement Learning

DGX agent

arXiv:2604.26516v1 Announce Type: cross Abstract: Offline reinforcement learning (RL) agents often fail when deployed, as the gap between training datasets and real environments leads to unsafe behavi

model-releasesarxiv-cs-ai
30 Apr 2026
Agents

MappingEvolve: LLM-Driven Code Evolution for Technology Mapping

DGX agent

arXiv:2604.26591v1 Announce Type: cross Abstract: Technology mapping is a critical yet challenging stage in logic synthesis. While Large Language Models (LLMs) have been applied to generate optimizati

agentsarxiv-cs-ai
30 Apr 2026
Model Releases

BitRL: Reinforcement Learning with 1-bit Quantized Language Models for Resource-Constrained Edge Deployment

DGX agent

arXiv:2604.24273v1 Announce Type: new Abstract: The deployment of intelligent reinforcement learning (RL) agents on resource-constrained edge devices remains a fundamental challenge due to the substan

model-releasesarxiv-cs-lg
28 Apr 2026
Model Releases

Don't Make the LLM Read the Graph: Make the Graph Think

DGX agent

arXiv:2604.23057v1 Announce Type: new Abstract: We investigate whether explicit belief graphs improve LLM performance in cooperative multi-agent reasoning. Through 3,000+ controlled trials across four

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

ESIA: An Energy-Based Spatiotemporal Interaction-Aware Framework for Pedestrian Intention Prediction

DGX agent

arXiv:2604.23728v1 Announce Type: cross Abstract: Recent advances in autonomous driving have motivated research on pedestrian intention prediction, which aims to infer future crossing decisions and ac

agentsarxiv-cs-ai
28 Apr 2026
Model Releases

LEGO: An LLM Skill-Based Front-End Design Generation Platform

DGX agent

arXiv:2604.23355v1 Announce Type: new Abstract: Existing LLM-based EDA agents are often isolated task-specific systems. This leads to repeated engineering effort and limited reuse of successful design

model-releasesarxiv-cs-ai
28 Apr 2026
Research

Leveraging Human Feedback for Semantically-Relevant Skill Discovery

DGX agent

arXiv:2604.24127v1 Announce Type: cross Abstract: Unsupervised skill discovery in reinforcement learning aims to intrinsically motivate agents to discover diverse and useful behaviours. However, uncon

researcharxiv-cs-ai
28 Apr 2026
Model Releases

NeuroClaw Technical Report

DGX agent

arXiv:2604.24696v1 Announce Type: new Abstract: Agentic artificial intelligence systems promise to accelerate scientific workflows, but neuroimaging poses unique challenges: heterogeneous modalities (

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

No Test Cases, No Problem: Distillation-Driven Code Generation for Scientific Workflows

DGX agent

arXiv:2604.23106v1 Announce Type: cross Abstract: Existing multi-agent Large Language Model (LLM) frameworks for code generation typically use execution feedback and improve iteratively using Input/Ou

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

On the Convergence of Jacobian-Free Backpropagation for Optimal Control Problems with Implicit Hamiltonians

DGX agent

arXiv:2602.00921v2 Announce Type: replace-cross Abstract: Optimal feedback control with implicit Hamiltonians poses a fundamental challenge for learning-based value function methods due to the absence

agentsarxiv-cs-lg
28 Apr 2026
Model Releases

Scheming Ability in LLM-to-LLM Strategic Interactions

DGX agent

arXiv:2510.12826v2 Announce Type: replace-cross Abstract: As large language model (LLM) agents are deployed autonomously in diverse contexts, evaluating their capacity for strategic deception becomes

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Cross-Stage Coherence in Hierarchical Driving VQA: Explicit Baselines and Learned Gated Context Projectors

DGX agent

arXiv:2604.22560v1 Announce Type: cross Abstract: Graph Visual Question Answering (GVQA) for autonomous driving organizes reasoning into ordered stages, namely Perception, Prediction, and Planning, wh

agentsarxiv-cs-ai
27 Apr 2026
Model Releases

Rethinking Publication: A Certification Framework for AI-Enabled Research

DGX agent

arXiv:2604.22026v1 Announce Type: new Abstract: AI research pipelines now produce a growing share of publishable academic output, including work that meets existing peer-review standards for quality a

model-releasesarxiv-cs-ai
27 Apr 2026
← Previous
1…149150151152153…236
Next →