AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

SubTGraph: Large-Scale Subterranean Environment Synthesis with Controllable Topological Variability for Robotic Autonomy Validation

DGX agent

arXiv:2605.20917v1 Announce Type: new Abstract: Subterranean (SubT) environments have been a frontier for autonomous robotics, driven by the push for automation of mining operations and the interest i

agentsarxiv-cs-ro
21 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Why Latent Actions Fail, and How to Prevent It

DGX agent

arXiv:2605.20223v1 Announce Type: new Abstract: Latent action models (LAMs) aim to learn action-like representations from unlabeled videos by compressing frame-to-frame changes. The frames of in-the-w

agentsarxiv-cs-cv
21 May 2026
Agents

A Geometric Analysis of Small-sized Language Model Hallucinations

DGX agent

arXiv:2602.14778v3 Announce Type: replace-cross Abstract: Hallucinations -- plausible but factually incorrect responses -- pose a major challenge to the reliability of Large Language Models (LLMs), es

agentsarxiv-cs-ai
20 May 2026
Agents

A Logistic Regression Model to Predict Malaria Severity in Children

DGX agent

arXiv:2605.18900v1 Announce Type: cross Abstract: One of the main causes of death around the globe is malaria. Researchers have sought to develop predictive models for malaria outbreaks based on meteo

agentsarxiv-cs-lg
20 May 2026
Agents

Causal Evidence that Language Models use Confidence to Drive Behavior

DGX agent

arXiv:2603.22161v2 Announce Type: replace Abstract: Metacognition -- assessing the quality of one's own cognitive performance -- guides adaptive behavior across species. Substantial research demonstra

agentsarxiv-cs-lg
20 May 2026
Agents

CLUE: Adaptively Prioritized Contextual Cues by Leveraging a Unified Semantic Map for Effective Zero-Shot Object-Goal Navigation

DGX agent

arXiv:2605.19206v1 Announce Type: new Abstract: Zero-shot object-goal navigation (ZSON) is a challenging problem in robotics that requires a comprehensive understanding of both language and visual obs

agentsarxiv-cs-ro
20 May 2026
Agents

DECOR: Auditing LLM Deception via Information Manipulation Theory

DGX agent

arXiv:2605.19270v1 Announce Type: new Abstract: Large language models can deceive by subtly manipulating truthful information -- omitting key facts, shifting focus, or obscuring meaning -- making such

agentsarxiv-cs-cl
20 May 2026
Safety

Dual-Gated Epistemic Time-Dilation: Autonomous Compute Modulation in Asynchronous MARL

DGX agent

arXiv:2603.23722v2 Announce Type: replace-cross Abstract: While Multi-Agent Reinforcement Learning (MARL) algorithms achieve unprecedented successes across complex continuous domains, their standard d

safetyarxiv-cs-lg
20 May 2026
Safety

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

DGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

safetyarxiv-cs-ai
20 May 2026
Agents

FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models

DGX agent

arXiv:2605.19111v1 Announce Type: cross Abstract: Existing text-to-image (T2I) evaluation metrics mainly assess whether generated images align with information explicitly stated in the prompt, but oft

agentsarxiv-cs-ai
20 May 2026
Agents

High-quality generation of dynamic game content via small language models: A proof of concept

DGX agent

arXiv:2601.23206v2 Announce Type: replace Abstract: Large language models (LLMs) offer promise for dynamic game content generation, but they face critical barriers, including narrative incoherence and

agentsarxiv-cs-ai
20 May 2026
Agents

Hybrid Training for Vision-Language-Action Models

DGX agent

arXiv:2510.00600v2 Announce Type: replace-cross Abstract: Using Large Language Models to produce intermediate thoughts, a.k.a. Chain-of-thought (CoT), before providing an answer has been a successful

agentsarxiv-cs-ai
20 May 2026
Agents

Library Drift: Diagnosing and Fixing a Silent Failure Mode in Self-Evolving LLM Skill Libraries

DGX agent

arXiv:2605.19576v1 Announce Type: new Abstract: Self-evolving skill libraries face a silent failure mode we term library drift: unbounded skill accumulation without outcome-driven lifecycle management

agentsarxiv-cs-ai
20 May 2026
Agents

Operationalising Artificial Intelligence Bills of Materials (AIBOMs) for Verifiable AI Provenance and Lifecycle Assurance

DGX agent

arXiv:2605.19755v1 Announce Type: cross Abstract: Artificial Intelligence (AI) systems are increasingly dependent on complex, multi-layered software supply chains that introduce challenges for reprodu

agentsarxiv-cs-ai
20 May 2026
Model Releases

Synthesis and Evaluation of Long-term History-aware Medical Dialogue

DGX agent

arXiv:2605.19766v1 Announce Type: cross Abstract: An effective healthcare agent must be able to recall and reason over a patient's longitudinal medical history. However, the absence of datasets with r

model-releasesarxiv-cs-ai
20 May 2026
Agents

The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection

DGX agent

arXiv:2605.18857v1 Announce Type: cross Abstract: For most of the history of information retrieval (IR), search results were designed for human consumers who could scan, filter, and discard irrelevant

agentsarxiv-cs-ai
20 May 2026
Agents

YAC: Bridging Natural Language and Interactive Visual Exploration with Generative AI for Biomedical Data Discovery

DGX agent

arXiv:2509.19182v2 Announce Type: replace-cross Abstract: Incorporating natural language input has the potential to improve the capabilities of biomedical data discovery interfaces. However, user inte

agentsarxiv-cs-ai
20 May 2026
Agents

A Mechanistic Model for Collective Motion from Sensorimotor Regularities

DGX agent

arXiv:2605.16522v1 Announce Type: new Abstract: Collective behavior in animals has long been modeled through self-propelled particle models, which reproduce striking group-level phenomena through abst

agentsarxiv-cs-ro
19 May 2026
Model Releases

A Pilot Benchmark for NL-to-FOL Translation in Planetary Exploration

DGX agent

arXiv:2605.17911v1 Announce Type: new Abstract: Future planetary exploration envisions autonomous robotic agents operating under severe communication constraints, without global positioning, and with

model-releasesarxiv-cs-cl
19 May 2026
Agents

Baba in Wonderland: Online Self-Supervised Dynamics Discovery for Executable World Models

DGX agent

arXiv:2605.16725v1 Announce Type: new Abstract: Executable world models can be read, edited, executed, and reused for planning, but only if the program captures the environment's transition law rather

agentsarxiv-cs-ai
19 May 2026
Model Releases

Causely: A Causal Intelligence Layer for Enterprise AI A Benchmark Study on SRE and Reliability Workflows

DGX agent

arXiv:2605.18327v1 Announce Type: new Abstract: AI agents deployed into SRE workflows currently derive their understanding of environment state from raw observability telemetry at query time, paying a

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

CooT: Learning to Coordinate In-Context with Coordination Transformers

DGX agent

arXiv:2506.23549v3 Announce Type: replace Abstract: Effective coordination among unfamiliar partners remains a major challenge in multi-agent systems. Existing approaches, such as population-based met

model-releasesarxiv-cs-ai
19 May 2026
Agents

DeepArrhythmia: Segment-Contextualized ECG Arrhythmia Classification via Selective Evidence Acquisition

DGX agent

arXiv:2605.16441v1 Announce Type: cross Abstract: Beat-level Electrocardiography (ECG) arrhythmia detection aims to assign an arrhythmia class to each beat in a recording, yet many existing systems tr

agentsarxiv-cs-ai
19 May 2026
Agents

Democratizing Large-Scale Re-Optimization with LLM-Guided Model Patches

DGX agent

arXiv:2605.18692v1 Announce Type: new Abstract: Optimization models developed by operations research (OR) experts are often deployed as decision-support systems in industrial settings. However, real-w

agentsarxiv-cs-ai
19 May 2026
Model Releases

DocReward: A Document Reward Model for Structuring and Stylizing

DGX agent

arXiv:2510.11391v3 Announce Type: replace-cross Abstract: Recent agentic workflows automate professional document generation but focus narrowly on textual quality, overlooking structural and stylistic

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

ESI-Bench: Towards Embodied Spatial Intelligence that Closes the Perception-Action Loop

DGX agent

arXiv:2605.18746v1 Announce Type: cross Abstract: Spatial intelligence unfolds through a perception-action loop: agents act to acquire observations, and reason about how observations vary as a functio

model-releasesarxiv-cs-ai
19 May 2026
Agents

Generative AI Advertising as a Problem of Trustworthy Commercial Intervention

DGX agent

arXiv:2605.18673v1 Announce Type: cross Abstract: Major deployed generative AI advertising systems preserve a visible boundary between commercial content and AI-generated responses. Yet empirical rese

agentsarxiv-cs-cl
19 May 2026
Agents

GraphMind: From Operational Traces to Self-Evolving Workflow Automation

DGX agent

arXiv:2605.17617v1 Announce Type: new Abstract: Complex operational workflows coordinating personnel, tools, and information are central to enterprise operations, yet end-to-end automation remains cha

agentsarxiv-cs-ai
19 May 2026
Model Releases

Hunt Instead of Wait: Evaluating Deep Data Research on Large Language Models

DGX agent

arXiv:2602.02039v2 Announce Type: replace Abstract: The agency expected of Agentic Large Language Models goes beyond answering correctly, requiring autonomy to set goals and decide what to explore. We

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

MARS: Technical Report for the CASTLE Challenge at EgoVis 2026

DGX agent

arXiv:2605.18176v1 Announce Type: cross Abstract: This report presents MARS, short for Multimodal Agentic Reasoning with Source selection, our system for the CASTLE Challenge at EgoVis 2026. Participa

model-releasesarxiv-cs-ai
19 May 2026
Agents

MATE: Solving Contextual Markov Decision Processes with Memory of Accumulated Transition Embeddings

DGX agent

arXiv:2605.17431v1 Announce Type: cross Abstract: We propose MATE, a simple yet effective memory architecture for solving Contextual Markov Decision Processes (CMDPs), a family of MDPs parameterized b

agentsarxiv-cs-ai
19 May 2026
Agents

Modelling Customer Trajectories with Reinforcement Learning for Practical Retail Insights

DGX agent

arXiv:2605.18449v1 Announce Type: cross Abstract: Understanding customer movement within retail spaces is essential for optimizing store layouts. Real-world trajectory data can provide highly accurate

agentsarxiv-cs-ai
19 May 2026
Model Releases

OpenJarvis: Personal AI, On Personal Devices

DGX agent

arXiv:2605.17172v1 Announce Type: cross Abstract: Personal AI stacks, like OpenClaw and Hermes Agent, are becoming central to daily work, yet they route nearly every query (often over sensitive local

model-releasesarxiv-cs-ai
19 May 2026
Model Releases

Prompts Don't Protect: Architectural Enforcement via MCP Proxy for LLM Tool Access Control

DGX agent

arXiv:2605.18414v1 Announce Type: cross Abstract: Large language models increasingly operate as autonomous agents that select and invoke tools from large registries. We identify a critical gap: when u

model-releasesarxiv-cs-ai
19 May 2026
Agents

Spatiotemporal Robustness of Temporal Logic Tasks using Multi-Objective Reasoning

DGX agent

arXiv:2603.29868v2 Announce Type: replace Abstract: The reliability of autonomous systems depends on their robustness, i.e., their ability to meet their objectives under uncertainty. In this paper, we

agentsarxiv-cs-ai
19 May 2026
Model Releases

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning

DGX agent

arXiv:2605.18109v1 Announce Type: new Abstract: In real home deployments, household agents must often operate from a complete household scene and a situated household request, rather than from a clean

model-releasesarxiv-cs-ai
19 May 2026
Agents

The Journal of Prompt-Engineered (Moral) Philosophy Or: Why AI-Assisted Ethics Research Requires Process Transparency

DGX agent

arXiv:2511.08639v3 Announce Type: replace-cross Abstract: Existing AI disclosure mandates in scholarship require that AI assistance be reported but leave transparency philosophically unspecified: they

agentsarxiv-cs-ai
19 May 2026
Agents

Transfer Learning for Customized Car Racing Environments

DGX agent

arXiv:2605.17928v1 Announce Type: cross Abstract: Transfer Learning, a technique where a model/agent can use the knowledge/expertise that it gained from one task and exploit that to solve another clos

agentsarxiv-cs-lg
19 May 2026
Agents

Attribute-Grounded Selective Reasoning for Artwork Emotion Understanding with Multimodal Large Language Models

DGX agent

arXiv:2605.15755v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can produce fluent artwork emotion explanations, but they often suffer from attribute flooding: they enumerate

agentsarxiv-cs-cv
18 May 2026
Model Releases

PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control

DGX agent

arXiv:2605.15963v1 Announce Type: new Abstract: Large vision-language models have significantly advanced GUI agents, enabling executable interaction across web, mobile, and desktop interfaces. Yet the

model-releasesarxiv-cs-ai
18 May 2026
Agents

Training on Documents About Monitoring Leads to CoT Obfuscation

DGX agent

arXiv:2605.15257v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is one of the most promising tools we have for detecting model misbehavior, but its effectiveness depends on models fa

agentsarxiv-cs-lg
18 May 2026
Agents

WorldAct: Activating Monolithic 3D Worlds into Interactive-Ready Object-Centric Scenes

DGX agent

arXiv:2605.15843v1 Announce Type: new Abstract: Recent 3D world modeling systems based on generative scene synthesis, such as Marble, can create coherent and explorable 3D environments, yet their outp

agentsarxiv-cs-cv
18 May 2026
Model Releases

Beyond AI as Assistants: Toward Autonomous Discovery in Cosmology

DGX agent

arXiv:2605.14791v1 Announce Type: cross Abstract: Recent advances in artificial intelligence (AI) agents are pushing AI beyond tools toward autonomous scientific discovery. We discuss two complementar

model-releasesarxiv-cs-ai
15 May 2026
Agents

Chrono-Gymnasium: An Open-Source, Gymnasium-Compatible Distributed Simulation Framework

DGX agent

arXiv:2605.14911v1 Announce Type: new Abstract: High-fidelity physics simulation is essential for closing the sim-to-real gap in robotics and complex mechanical systems. However, the computational ove

agentsarxiv-cs-ro
15 May 2026
Safety

Collaborative Yet Personalized Policy Training: Single-Timescale Federated Actor-Critic

DGX agent

arXiv:2605.14423v1 Announce Type: cross Abstract: Despite the popularity of the actor-critic method and the practical needs of collaborative policy training, existing works typically either overlook e

safetyarxiv-cs-ai
15 May 2026
Model Releases

Deceive, Detect, and Disclose: Large Language Models Play Mini-Mafia

DGX agent

arXiv:2509.23023v3 Announce Type: replace Abstract: Large language models are increasingly deployed in multi-agent settings whose outcomes hinge on social intelligence, motivating evaluations of their

model-releasesarxiv-cs-ai
15 May 2026
Agents

From Plans to Pixels: Learning to Plan and Orchestrate for Open-Ended Image Editing

DGX agent

arXiv:2605.15181v1 Announce Type: new Abstract: Modern image editing models produce realistic results but struggle with abstract, multi step instructions (e.g., ``make this advertisement more vegetari

agentsarxiv-cs-cv
15 May 2026
Agents

Fully Dynamic Rebalancing in Dockless Bike-Sharing Systems via Deep Reinforcement Learning

DGX agent

arXiv:2605.14501v1 Announce Type: cross Abstract: This paper proposes a fully dynamic Deep Reinforcement Learning (DRL) method for rebalancing dockless bike-sharing systems, overcoming the limitations

agentsarxiv-cs-ai
15 May 2026
← Previous
1…147148149150151…236
Next →