AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

Deterministic Event-Graph Substrates as World Models for Counterfactual Reasoning

DGX agent

arXiv:2605.15967v1 Announce Type: new Abstract: We study event-graph substrates: a class of world models that represent agent state as an append-only log of typed RDF triples and answer counterfactual

model-releasesarxiv-cs-ai
18 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding

DGX agent

arXiv:2605.15542v1 Announce Type: new Abstract: GUI agents powered by Multimodal Large Language Models (MLLMs) have demonstrated impressive capability in understanding and executing user instructions.

researcharxiv-cs-ai
18 May 2026
Model Releases

Hybrid LLM-based Intelligent Framework for Robot Task Scheduling

DGX agent

arXiv:2605.15486v1 Announce Type: cross Abstract: This study introduces intelligent frameworks that use Large Language Models (LLMs) to improve task scheduling for construction robots. The LLM is fed

model-releasesarxiv-cs-ai
18 May 2026
Model Releases

Improved Bounds for Reward-Agnostic and Reward-Free Exploration

DGX agent

arXiv:2602.16363v2 Announce Type: replace Abstract: We study reward-free and reward-agnostic exploration in episodic finite-horizon Markov decision processes (MDPs), where an agent explores an unknown

model-releasesarxiv-cs-lg
18 May 2026
Model Releases

Large Language Models as Optimization Controllers: Adaptive Continuation for SIMP Topology Optimization

DGX agent

arXiv:2603.25099v2 Announce Type: replace-cross Abstract: We present a framework in which a large language model (LLM) acts as an online adaptive controller for SIMP topology optimization, replacing c

model-releasesarxiv-cs-ai
18 May 2026
Safety

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning

DGX agent

arXiv:2605.15975v1 Announce Type: new Abstract: We tackle the challenge of building embodied AI agents that can reliably solve long-horizon planning problems. Imitation learning from demonstrations ha

safetyarxiv-cs-ai
18 May 2026
Model Releases

Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding

DGX agent

arXiv:2605.15342v1 Announce Type: new Abstract: Video reasoning models are a core component of egocentric and embodied agents. However, standard benchmarks for assessing models provide only evaluation

model-releasesarxiv-cs-cv
18 May 2026
Safety

ScreenSearch: Uncertainty-Aware OS Exploration

DGX agent

arXiv:2605.16024v1 Announce Type: new Abstract: Desktop GUI agents operate under partial observability: visually similar screens can correspond to different underlying workflow states, so locally plau

safetyarxiv-cs-ai
18 May 2026
Research

BOOKMARKS: Efficient Active Storyline Memory for Role-playing

DGX agent

arXiv:2605.14169v1 Announce Type: new Abstract: Memory systems are critical for role-playing agents (RPAs) to maintain long-horizon consistency. However, existing RPA memory methods (e.g., profiling)

researcharxiv-cs-cl
15 May 2026
Research

EARL: Towards a Unified Analysis-Guided Reinforcement Learning Framework for Egocentric Interaction Reasoning and Pixel Grounding

DGX agent

arXiv:2605.14742v1 Announce Type: new Abstract: Understanding human--environment interactions from egocentric vision is essential for assistive robotics and embodied intelligent agents, yet existing m

researcharxiv-cs-cv
15 May 2026
Hardware

Falkor-IRAC: Graph-Constrained Generation for Verified Legal Reasoning in Indian Judicial AI

DGX agent

arXiv:2605.14665v1 Announce Type: new Abstract: Legal reasoning is not semantic similarity search. A court judgment encodes constrained symbolic reasoning: precedent propagation, procedural state tran

hardwarearxiv-cs-ai
15 May 2026
Model Releases

MemLens: Benchmarking Multimodal Long-Term Memory in Large Vision-Language Models

DGX agent

arXiv:2605.14906v1 Announce Type: new Abstract: Memory is essential for large vision-language models (LVLMs) to handle long, multimodal interactions, with two method directions providing this capabili

model-releasesarxiv-cs-cv
15 May 2026
Local Ai

Self-Pruned Key-Value Attention: Learning When to Write by Predicting Future Utility

DGX agent

arXiv:2605.14037v1 Announce Type: cross Abstract: Under modern test-time compute and agentic paradigms, language models process ever-longer sequences. Efficient text generation with transformer archit

local-aiarxiv-cs-cl
15 May 2026
Safety

SpeakerLLM: A Speaker-Specialized Audio-LLM for Speaker Understanding and Verification Reasoning

DGX agent

arXiv:2605.15044v1 Announce Type: cross Abstract: As audio-first agents become increasingly common in physical AI, conversational robots, and screenless wearables, audio large language models (audio-L

safetyarxiv-cs-ai
15 May 2026
Model Releases

SPIN: Structural LLM Planning via Iterative Navigation for Industrial Tasks

DGX agent

arXiv:2605.14051v1 Announce Type: new Abstract: Industrial LLM agent systems often separate planning from execution, yet LLM planners frequently produce structurally invalid or unnecessarily long work

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

VerbalValue: A Socially Intelligent Virtual Host for Sales-Driven Live Commerce

DGX agent

arXiv:2605.14542v1 Announce Type: new Abstract: A skilled live-commerce host is not merely a narrator, but a sales agent who converts viewer curiosity into purchase intent through expert product knowl

model-releasesarxiv-cs-ai
15 May 2026
Safety

Causality-Aware End-to-End Autonomous Driving via Ego-Centric Joint Scene Modeling

DGX agent

arXiv:2605.13646v1 Announce Type: cross Abstract: End-to-end autonomous driving, which bypasses traditional modular pipelines by directly predicting future trajectories from sensor inputs, has recentl

safetyarxiv-cs-ai
14 May 2026
Safety

Differentiable Evolutionary Reinforcement Learning

DGX agent

arXiv:2512.13399v2 Announce Type: replace Abstract: Crafting effective reward signals remains a central challenge in Reinforcement Learning (RL), especially for complex reasoning tasks. Existing autom

safetyarxiv-cs-ai
14 May 2026
Applications

EgoForce: Robust Online Egocentric Motion Reconstruction via Diffusion Forcing

DGX agent

arXiv:2605.13041v1 Announce Type: new Abstract: With recent advances in embodied agents and AR devices, egocentric observations are readily available as input for real-world interactive online applica

applicationsarxiv-cs-cv
14 May 2026
Model Releases

Embodied Neurocomputation: A Framework for Interfacing Biological Neural Cultures with Scaled Task-Driven Validation

DGX agent

arXiv:2605.13315v1 Announce Type: cross Abstract: Biological neural networks (BNNs) have been established as a powerful and adaptive substrate that offer the potential for incredibly energy and data e

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Language Model Goal Selection Differs from Humans' in a Self-Directed Learning Task

DGX agent

arXiv:2603.03295v2 Announce Type: replace-cross Abstract: Whether in agentic workflows, social studies, or chat settings, large language models (LLMs) are increasingly being asked to replace humans in

model-releasesarxiv-cs-ai
14 May 2026
Safety

Learning When to Act: Communication-Efficient Reinforcement Learning via Run-Time Assurance

DGX agent

arXiv:2605.12561v1 Announce Type: new Abstract: Safe reinforcement learning (RL) typically asks extit{what} an agent should do. We ask extit{when} it needs to act, and show that a single policy can jo

safetyarxiv-cs-lg
14 May 2026
Tutorials

Probabilistic Prediction Markets with Intermittent Contributions

DGX agent

arXiv:2510.13385v3 Announce Type: replace Abstract: Although both data availability and the demand for accurate forecasts are increasing, collaboration between stakeholders is often constrained by dat

tutorialsarxiv-cs-lg
14 May 2026
Model Releases

ToolWeave: Structured Synthesis of Complex Multi-Turn Tool-Calling Dialogues

DGX agent

arXiv:2605.12521v1 Announce Type: cross Abstract: Multi-turn tool calling is essential for LLMs to function as autonomous agents, yet synthesizing the training data required for these capabilities rem

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

3D-Belief: Embodied Belief Inference via Generative 3D World Modeling

DGX agent

arXiv:2605.11367v1 Announce Type: new Abstract: Recent advances in visual generative models have highlighted the promise of learning generative world models. However, most existing approaches frame wo

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Covering Human Action Space for Computer Use: Data Synthesis and Benchmark

DGX agent

arXiv:2605.12501v1 Announce Type: new Abstract: Computer-use agents (CUAs) automate on-screen work, as illustrated by GPT-5.4 and Claude. Yet their reliability on complex, low-frequency interactions i

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

DarkQA: Benchmarking Vision-Language Models on Visual-Primitive Question Answering in Low-Light Indoor Scenes

DGX agent

arXiv:2512.24985v4 Announce Type: replace Abstract: Vision Language Models (VLMs) are increasingly adopted as central reasoning modules for embodied agents. Existing benchmarks evaluate their capabili

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Efficient Algorithms for Logistic Contextual Slate Bandits with Bandit Feedback

DGX agent

arXiv:2506.13163v3 Announce Type: replace Abstract: We study the Logistic Contextual Slate Bandit problem, where, at each round, an agent selects a slate of N items from an exponentially large set (of

model-releasesarxiv-cs-lg
13 May 2026
Model Releases

Fine-Tuning Large Language Models for Cooperative Tactical Deconfliction of Small Unmanned Aerial Systems

DGX agent

arXiv:2603.28561v2 Announce Type: replace Abstract: The growing deployment of small Unmanned Aerial Systems (sUASs) in low-altitude airspaces has increased the need for reliable tactical deconfliction

model-releasesarxiv-cs-ro
13 May 2026
Model Releases

Nautilus: From One Prompt to Plug-and-Play Robot Learning

DGX agent

arXiv:2605.11665v1 Announce Type: new Abstract: Robot learning research is fragmented across policy families, benchmark suites, and real robots; each implementation is entangled with the others in a c

model-releasesarxiv-cs-ro
13 May 2026
Safety

PriorZero: Bridging Language Priors and World Models for Decision Making

DGX agent

arXiv:2605.12289v1 Announce Type: new Abstract: Leveraging the rich world knowledge of Large Language Models (LLMs) to enhance Reinforcement Learning (RL) agents offers a promising path toward general

safetyarxiv-cs-lg
13 May 2026
Research

Shapley Value Approximation Based on k-Additive Games

DGX agent

arXiv:2502.04763v2 Announce Type: replace-cross Abstract: The Shapley value is the prevalent solution for fair division problems in which a payout is to be divided among multiple agents. By adopting a

researcharxiv-cs-lg
13 May 2026
Safety

Transferable Delay-Aware Reinforcement Learning via Implicit Causal Graph Modeling

DGX agent

arXiv:2605.12312v1 Announce Type: new Abstract: Random delays weaken the temporal correspondence between actions and subsequent state feedback, making it difficult for agents to identify the true prop

safetyarxiv-cs-lg
13 May 2026
Hardware

TriBand-BEV: Real-Time LiDAR-Only 3D Pedestrian Detection via Height-Aware BEV and High-Resolution Feature Fusion

DGX agent

arXiv:2605.12220v1 Announce Type: new Abstract: Safe autonomous agents and mobile robots need fast real time 3D perception, especially for vulnerable road users (VRUs) such as pedestrians. We introduc

hardwarearxiv-cs-cv
13 May 2026
Model Releases

UHR-Micro: Diagnosing and Mitigating the Resolution Illusion in Earth Observation VLMs

DGX agent

arXiv:2605.12237v1 Announce Type: new Abstract: Vision-Language Models (VLMs) increasingly operate on ultra-high-resolution (UHR) Earth observation imagery, yet they remain vulnerable to a severe scal

model-releasesarxiv-cs-cv
13 May 2026
Model Releases

Beyond Self-Play and Scale: A Behavior Benchmark for Generalization in Autonomous Driving

DGX agent

arXiv:2605.10034v1 Announce Type: new Abstract: Recent Autonomous Driving (AD) works such as GigaFlow and PufferDrive have unlocked Reinforcement Learning (RL) at scale as a training strategy for driv

model-releasesarxiv-cs-ro
12 May 2026
Research

Beyond Thinking: Imagining in 360^irc for Humanoid Visual Search

DGX agent

arXiv:2605.09146v1 Announce Type: new Abstract: Humanoid Visual Search (HVS) requires agents to actively explore immersive 360^irc environments. While prior methods treat this as a monolithic task rel

researcharxiv-cs-cv
12 May 2026
Model Releases

Do Benchmarks Underestimate LLM Performance? Evaluating Hallucination Detection With LLM-First Human-Adjudicated Assessment

DGX agent

arXiv:2605.08462v1 Announce Type: cross Abstract: Hallucination remains a persistent challenge in Large Language Models (LLMs), particularly in context-grounded settings such as RAG and agentic AI sys

model-releasesarxiv-cs-ai
12 May 2026
Tutorials

Efficient Estimation of Kernel Surrogate Models for Task Attribution

DGX agent

arXiv:2602.03783v2 Announce Type: replace-cross Abstract: Modern AI agents such as large language models are trained on diverse tasks -- translation, code generation, mathematical reasoning, and text

tutorialsarxiv-cs-ai
12 May 2026
Safety

Equivariant Reinforcement Learning for Clifford Quantum Circuit Synthesis

DGX agent

arXiv:2605.10910v1 Announce Type: cross Abstract: We consider the problem of synthesizing Clifford quantum circuits for devices with all-to-all qubit connectivity. We approach this task as a reinforce

safetyarxiv-cs-lg
12 May 2026
Safety

Execution Envelopes: A Shared Admission Contract for Backend AI Execution Requests

DGX agent

arXiv:2605.08267v1 Announce Type: cross Abstract: Enterprise AI backends increasingly admit heterogeneous execution requests across model deployment, inference, evaluation, data movement, and agentic

safetyarxiv-cs-ai
12 May 2026
Safety

Governed Metaprogramming for Intelligent Systems: Reclassifying Eval as a Governed Effect

DGX agent

arXiv:2605.05248v2 Announce Type: replace-cross Abstract: AI systems increasingly synthesize executable structure at runtime: LLMs generate programs, agents construct workflows,self-improving systems

safetyarxiv-cs-ai
12 May 2026
Model Releases

Higher Resolution, Better Generalization: Unlocking Visual Scaling in Deep Reinforcement Learning

DGX agent

arXiv:2605.10546v1 Announce Type: new Abstract: Pixel-based deep reinforcement learning agents are typically trained on heavily downsampled visual observations, a convention inherited from early bench

model-releasesarxiv-cs-lg
12 May 2026
Research

Learning Strategic Value and Cooperation in Multi-Player Stochastic Games through Side Payments

DGX agent

arXiv:2303.05307v2 Announce Type: replace-cross Abstract: We study general-sum, multi-player stochastic games with transferable utility, motivated by settings where agents can use side payments to mak

researcharxiv-cs-ai
12 May 2026
Safety

LLM Advertisement based on Neuron Auctions

DGX agent

arXiv:2605.08326v1 Announce Type: cross Abstract: As Large Language Models (LLMs) transition into conversational agents, generative advertising emerges as a crucial monetization strategy. However, emb

safetyarxiv-cs-ai
12 May 2026
Model Releases

MLS-Bench: A Holistic and Rigorous Assessment of AI Systems on Building Better AI

DGX agent

arXiv:2605.08678v1 Announce Type: new Abstract: Modern AI progress has been driven by ML methods that are generalizable across settings and scalable to larger regimes. As large language models demonst

model-releasesarxiv-cs-lg
12 May 2026
Safety

Neural Co-state Policies: Structuring Hidden States in Recurrent Reinforcement Learning

DGX agent

arXiv:2605.05373v2 Announce Type: replace Abstract: A key capability of intelligent agents is operating under partial observability: reasoning and acting effectively despite missing or incomplete stat

safetyarxiv-cs-lg
12 May 2026
Model Releases

PDEAgent-Bench: A Multi-Metric, Multi-Library Benchmark for PDE Solver Generation

DGX agent

arXiv:2605.09636v1 Announce Type: new Abstract: PDE-to-solver code generation aims to automatically synthesize executable numerical solvers from partial differential equation (PDE) specifications. Thi

model-releasesarxiv-cs-ai
12 May 2026
← Previous
1…180181182183184…233
Next →