AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

LLM-Guided ANN Index Optimization for Human-Object Interaction Retrieval

DGX agent

arXiv:2606.05489v1 Announce Type: new Abstract: Retrieval systems underpin modern AI applications -- spanning visual search, recommendation engines, and multi-modal question answering. Modern multi-st

model-releasesarxiv-cs-cv
5 Jun 2026
Research
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

PHUMA: Physically Reliable Humanoid Locomotion Dataset

DGX agent

arXiv:2510.26236v2 Announce Type: replace Abstract: Motion imitation is a promising approach for humanoid locomotion, enabling agents to acquire humanlike behaviors. Existing methods typically rely on

researcharxiv-cs-ro
5 Jun 2026
Safety

QueryAgent-R1: Bridging Query Generation and Product Retrieval for E-Commerce Query Recommendation

DGX agent

arXiv:2606.05671v1 Announce Type: new Abstract: Query recommendation in e-commerce search aims to proactively suggest queries that match users' potential interests. However, existing methods mainly op

safetyarxiv-cs-cl
5 Jun 2026
Safety

RiskFlow: Fast and Faithful Safety-Critical Traffic Scenario Generation

DGX agent

arXiv:2606.06423v1 Announce Type: new Abstract: Safety-critical traffic scenario generation is essential for evaluating autonomous driving systems under rare but high-risk interactions. Existing diffu

safetyarxiv-cs-ro
5 Jun 2026
Safety

Adaptive Information Control for Search-Augmented LLM Reasoning

DGX agent

arXiv:2602.01672v2 Announce Type: replace Abstract: Search-augmented reasoning agents interleave multi-step reasoning with external retrieval, but uncontrolled retrieval can introduce redundant eviden

safetyarxiv-cs-cl
4 Jun 2026
Model Releases

BioBlue: Systematic runaway-optimiser-like LLM failure modes on biologically and economically aligned AI safety benchmarks for LLMs with simplified observation format

DGX agent

arXiv:2509.02655v3 Announce Type: replace-cross Abstract: Many AI alignment discussions of 'runaway optimisation' focus on RL agents: unbounded utility maximisers that over-optimise a proxy objective

model-releasesarxiv-cs-ai
4 Jun 2026
Applications

CYGNET: Cypher Gate for Neural Execution Triage and Cost Containment

DGX agent

arXiv:2606.04645v1 Announce Type: new Abstract: Language models acting as agents over knowledge graphs generate Cypher queries that fail structurally (crashing at the database) or semantically (execut

applicationsarxiv-cs-cl
4 Jun 2026
Safety

Policy Gradient for Continuous-Time Robust Markov Decision Processes

DGX agent

arXiv:2606.04335v1 Announce Type: new Abstract: The framework of robust Markov decision processes (RMDPs) allows the design of reinforcement learning agents that satisfy performance guarantees under w

safetyarxiv-cs-lg
4 Jun 2026
Local Ai

R-APS: Compositional Reasoning and In-Context Meta-Learning for Constrained Design via Reflective Adversarial Pareto Search

DGX agent

arXiv:2606.04823v1 Announce Type: new Abstract: Large language models (LLMs) are fluent on open-ended tasks, yet in agentic settings, where a system must plan, use tools, and act over extended horizon

local-aiarxiv-cs-ai
4 Jun 2026
Safety

Unlocking Proactivity in Task-Oriented Dialogue

DGX agent

arXiv:2605.22240v2 Announce Type: replace Abstract: Proactive task-oriented dialogue (TOD), such as outbound sales, demands a persuasive agent that actively probes the user's concerns and steers the c

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

Benchmarking Visual State Tracking in Multimodal Video Understanding

DGX agent

arXiv:2606.03920v1 Announce Type: new Abstract: Understanding a video requires more than recognizing isolated moments, as humans continuously track entities, states, and events over time. This capacit

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

Chatbots Output Meaningful (but Problematic) Language

DGX agent

arXiv:2606.02973v1 Announce Type: new Abstract: Are utterances by AI chatbots meaningful? Concretely, if a user asks, say, Anthropic's agent Claude, 'What is the capital of Spain?' and Claude answers,

model-releasesarxiv-cs-cl
3 Jun 2026
Model Releases

GN0: Toward a Unified Paradigm for Generation, Evaluation, and Policy Learning in Visual-Language Navigation

DGX agent

arXiv:2606.03682v1 Announce Type: new Abstract: Embodied navigation connects intelligent agents with the physical world and is fundamental for general robotic intelligence. Limited availability and qu

model-releasesarxiv-cs-ro
3 Jun 2026
Model Releases

MemoGen: Can Past Experience Improve Future Text-to-Image Generation?

DGX agent

arXiv:2606.03243v1 Announce Type: new Abstract: Modern text-to-image models have achieved strong visual synthesis, yet remain unreliable when prompts require implicit visual constraints, relational re

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

OVO-S-Bench: A Hierarchical Benchmark for Streaming Spatial Intelligence in Multimodal LLMs

DGX agent

arXiv:2606.03890v1 Announce Type: new Abstract: Multimodal agents in robotics, AR, and autonomous driving must reason about places and layouts from continuous egocentric streams, often using evidence

model-releasesarxiv-cs-cv
3 Jun 2026
Model Releases

See, Infer, Intervene: Proactive World Modeling for Goal-Oriented Social Intelligence

DGX agent

arXiv:2606.03371v1 Announce Type: new Abstract: Multimodal retail agents should not only recognize what a customer is doing, but also decide whether and how to assist before an explicit request is mad

model-releasesarxiv-cs-cl
3 Jun 2026
Applications

Solipsistic Superintelligence is Unlikely to be Cooperative

DGX agent

arXiv:2606.03237v1 Announce Type: new Abstract: AI's central challenge is shifting from capability to coexistence. The dominant paradigm in AI research focuses on developing powerful agents that treat

applicationsarxiv-cs-ai
3 Jun 2026
Model Releases

VistaHop: Benchmarking Multi-hop Visual Reasoning for Visual DeepSearch

DGX agent

arXiv:2606.03273v1 Announce Type: cross Abstract: Visual DeepSearch requires multimodal large reasoning model (MLRM) agents to answer complex visual queries by repeatedly inspecting image regions, gro

model-releasesarxiv-cs-ai
3 Jun 2026
Research

Bridging the 2D-3D Gap: A Hierarchical Semantic-Geometric Map for Vision Language Navigation

DGX agent

arXiv:2606.00095v1 Announce Type: cross Abstract: Vision-Language Navigation (VLN) enables embodied agents to reach target locations in unseen environments by following language instructions. Despite

researcharxiv-cs-ai
2 Jun 2026
Model Releases

Completion at the Boundary (CaB): Deployable Switching with Completion-Aware Control under Limited Calibration

DGX agent

arXiv:2606.00145v1 Announce Type: cross Abstract: Vision-language-action (VLA) agents can execute natural-language instructions, yet deployed systems still lack an operational interface: deciding when

model-releasesarxiv-cs-ai
2 Jun 2026
Research

Differing Roles of Leisure and Productivity in GDP - A Machine Learning based comparative analysis of Germany and USA

DGX agent

arXiv:2606.01234v1 Announce Type: cross Abstract: The GDP of a country is modelled as the relative interaction between two agents - working hours, reflecting the social choice of a population, and Tot

researcharxiv-cs-cv
2 Jun 2026
Model Releases

Ego-METAS: Egocentric online Multimodal Energy-efficient Temporal Action Segmentation benchmark

DGX agent

arXiv:2606.02246v1 Announce Type: new Abstract: To operate in the physical world, embodied agents must perceive their environment in an 'always-on' fashion, selectively accessing the most informative

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Emergence of Exploration in Policy Gradient Reinforcement Learning via Retrying

DGX agent

arXiv:2606.00151v1 Announce Type: cross Abstract: In reinforcement learning (RL), agents benefit from exploration only because they repeatedly encounter similar states: trying different actions can im

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

HERO'S JOURNEY: Testing Complex Rule Induction with Text Games

DGX agent

arXiv:2606.02556v1 Announce Type: new Abstract: We introduce HERO'S JOURNEY, a benchmark for rule induction in goal-directed episodic tasks, where agents must infer hidden rules from demonstrations an

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

Hierarchical Semantic-Augmented Navigation: Optimal Transport and Graph-Driven Reasoning for Vision-Language Navigation

DGX agent

arXiv:2606.01565v1 Announce Type: cross Abstract: Vision-Language Navigation in Continuous Environments (VLN-CE) poses a formidable challenge for autonomous agents, requiring seamless integration of n

safetyarxiv-cs-cv
2 Jun 2026
Safety

KISS: Keeping it Simple and Slotted when Learning to Communicate over Wireless

DGX agent

arXiv:2606.00266v1 Announce Type: cross Abstract: A long-standing challenge in distributed wireless systems is ensuring efficient and fair random channel access. Existing solutions often address speci

safetyarxiv-cs-lg
2 Jun 2026
Research

Learning Action-Conditional and Object-Centric Gaussian Splatting World Models for Rigid Objects

DGX agent

arXiv:2606.01950v1 Announce Type: cross Abstract: World models enable intelligent agents to predict the consequences of their actions on the environment. In this paper, we propose Multi Rigid Object G

researcharxiv-cs-cv
2 Jun 2026
Model Releases

MASER: Modality-Adaptive Specialist Routing for Embodied 3D Spatial Intelligence

DGX agent

arXiv:2606.02463v1 Announce Type: cross Abstract: In 3D environments, Embodied Agents answer spatially relevant questions through reasoning from a mixture of modalities including natural language, RGB

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

MindZero: Learning Online Mental Reasoning With Zero Annotations

DGX agent

arXiv:2606.00240v1 Announce Type: new Abstract: Effective real-world assistance requires AI agents with robust Theory of Mind (ToM): inferring human mental states from their behavior. Despite recent a

applicationsarxiv-cs-ai
2 Jun 2026
Research

PSG-Nav: Probabilistic Scene Graph Navigation via Multiverse Decision Making

DGX agent

arXiv:2606.01313v1 Announce Type: cross Abstract: Open-vocabulary navigation requires embodied agents to manage significant perception uncertainty stemming from semantic ambiguity and model errors. Ho

researcharxiv-cs-ai
2 Jun 2026
Safety

Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)

DGX agent

arXiv:2512.18333v2 Announce Type: replace-cross Abstract: This paper proposes a new Reinforcement Learning (RL) based control architecture for quadrotors. With the literature focusing on controlling t

safetyarxiv-cs-ai
2 Jun 2026
Safety

Robust Shielding for Safe Reinforcement Learning

DGX agent

arXiv:2606.00270v1 Announce Type: new Abstract: Shielding is an effective approach to formally guarantee the safety of reinforcement learning agents in Markov decision processes (MDPs). However, exist

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Simple Recipe Works: Vision-Language-Action Models are Natural Continual Learners with Reinforcement Learning

DGX agent

arXiv:2603.11653v2 Announce Type: replace Abstract: Continual Reinforcement Learning (CRL) for Vision-Language-Action (VLA) models is a promising direction toward self-improving embodied agents that c

model-releasesarxiv-cs-lg
2 Jun 2026
Hardware

STaR-KV: Spatio-Temporal Adaptive Re-weighting for KV Cache Compression in GUI Vision-Language Models

DGX agent

arXiv:2606.01790v1 Announce Type: cross Abstract: Vision-language-model-based graphical user interface (GUI) agents have shown broad automation capabilities, yet deployment is bottlenecked by a key-va

hardwarearxiv-cs-ai
2 Jun 2026
Model Releases

StreamingVLM: Real-Time Understanding for Infinite Video Streams

DGX agent

arXiv:2510.09608v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) could power real-time assistants and autonomous agents, but they face a critical challenge: understanding near-i

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

Trust Region On-Policy Distillation

DGX agent

arXiv:2606.01249v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) is a fundamental technique for efficient post-training of large language models (LLMs), with broad applications in agent

safetyarxiv-cs-cl
2 Jun 2026
Hardware

Deterministic Inference across Tensor Parallel Sizes That Eliminates Training-Inference Mismatch

DGX agent

arXiv:2511.17826v2 Announce Type: replace-cross Abstract: Deterministic inference is increasingly critical for large language model (LLM) applications such as LLM-as-a-judge evaluation, multi-agent sy

hardwarearxiv-cs-cl
1 Jun 2026
Model Releases

EGOSTREAM: A Diagnostic Benchmark for Streaming Episodic Memory in Egocentric Vision

DGX agent

arXiv:2605.31557v1 Announce Type: new Abstract: Continuous episodic memory is a core capability for autonomous agents operating in dynamic, real-world environments, yet current streaming video benchma

model-releasesarxiv-cs-cv
1 Jun 2026
Local Ai

GPU Forecasters: Language Models as Selective Surrogates for Kernel Runtime Optimization

DGX agent

arXiv:2605.31464v1 Announce Type: cross Abstract: GPU kernels are the workhorse of modern deep learning, and optimizing them (via evolutionary search or coding agents) usually requires repeated measur

local-aiarxiv-cs-ai
1 Jun 2026
Model Releases

LangMap: A Human-Verified Benchmark for Hierarchical Open-Vocabulary Goal Navigation

DGX agent

arXiv:2602.02220v2 Announce Type: replace Abstract: Language-conditioned goal navigation (LGN) requires agents to locate user-specified targets without step-by-step guidance. However, existing benchma

model-releasesarxiv-cs-cv
1 Jun 2026
Model Releases

Learning Randomized Reductions

DGX agent

arXiv:2412.18134v4 Announce Type: replace Abstract: Randomized self-reductions (RSRs) express f(x) using f evaluated at random correlated points, enabling self-correcting programs, instance-hiding pro

model-releasesarxiv-cs-lg
1 Jun 2026
Model Releases

Memory-Bound but Not Bandwidth-Limited: The Physical AI Inference Gap in Batch-1 LLM Decode

DGX agent

arXiv:2605.30571v1 Announce Type: cross Abstract: Physical AI systems, including robots, autonomous vehicles, embodied agents and edge copilots, often run a different inference workload from cloud LLM

model-releasesarxiv-cs-ai
1 Jun 2026
Model Releases

MLIPilot: LLM-Driven Auto-Research for Machine-Learned Interatomic Potentials

DGX agent

arXiv:2605.30889v1 Announce Type: cross Abstract: Constructing production-quality machine-learned interatomic potentials (MLIPs) requires balancing accuracy, dynamical stability, and computational thr

model-releasesarxiv-cs-lg
1 Jun 2026
Research

Physics-informed Goal-Conditioned Reinforcement Learning under Hybrid Contact Dynamics

DGX agent

arXiv:2605.30503v1 Announce Type: new Abstract: Learning to reach arbitrary goals from sparse feedback requires agents to infer a rich notion of reachability across state--goal pairs. Goal-conditioned

researcharxiv-cs-ro
1 Jun 2026
Safety

Preference-Aware Rubric Learning for Personalized Evaluation

DGX agent

arXiv:2605.31545v1 Announce Type: new Abstract: As Large Language Models (LLMs) evolve from general-purpose assistants to user-centric agents, personalization has become central to aligning model beha

safetyarxiv-cs-cl
1 Jun 2026
Safety

Structure-Induced Information for Rerooting Levin Tree Search

DGX agent

arXiv:2605.30664v1 Announce Type: new Abstract: Subgoal-based policy tree search, which uses a policy to guide search, is effective for complex single-agent deterministic problems but often relies on

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

Beyond Recall: Behavioral Specification as an Interpretive Layer for AI Personalization

DGX agent

arXiv:2605.28969v1 Announce Type: cross Abstract: If an AI agent makes decisions on a person's behalf, those decisions must align with its user. We introduce representational accuracy to measure how f

model-releasesarxiv-cs-ai
29 May 2026
Research

CoHyDE: Iterative Co-Training of LLM Rewriter & Dense Encoder for Tool Retrieval

DGX agent

arXiv:2605.29271v1 Announce Type: new Abstract: Tool retrieval over large API catalogs is a core bottleneck for LLM agents: user queries arrive in colloquial, often underspecified language, while the

researcharxiv-cs-ai
29 May 2026
← Previous
1…177178179180181…233
Next →