AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

Sampling-Based Coordination-Informed Multi-Objective Multi-Robot Reinforcement Learning

DGX agent

arXiv:2606.30893v1 Announce Type: new Abstract: Multi-robot systems must simultaneously optimize competing objectives while maintaining coordinated behavior. Existing multi-agent reinforcement learnin

safetyarxiv-cs-ro
1 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

The HydroGym Reinforcement Learning Platform for Fluid Dynamics

DGX agent

arXiv:2512.17534v2 Announce Type: replace-cross Abstract: Modeling and controlling fluids is critical across science and engineering. Effective flow control can increase lift, reduce drag, enhance mix

researcharxiv-cs-ai
1 Jul 2026
Safety

The Past Is Prologue: A Plug-in Controller for Selective Updates in Sequentially Evolving LLM Memory

DGX agent

arXiv:2606.31121v1 Announce Type: new Abstract: Sequentially evolving LLM memory enables agents to reuse past experience, but existing systems usually deploy each locally generated memory update witho

safetyarxiv-cs-ai
1 Jul 2026
Research

Who Determines the Meaning of an Emotion? Affective Sovereignty as an Epistemic Consequence of Measurement Limits

DGX agent

arXiv:2606.31442v1 Announce Type: new Abstract: Emotion-sensing AI is rapidly becoming embedded in vehicles, home appliances, dialogue agents, and social infrastructure, giving rise to a sphere in whi

researcharxiv-cs-ai
1 Jul 2026
Model Releases

Accelerating scientific discovery with Co-Scientist

DGX agent

arXiv:2502.18864v2 Announce Type: replace Abstract: Scientific discovery is driven by scientists generating novel hypotheses for complex problems that undergo rigorous experimental validation. To augm

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

CLOSER-VLN: Closed-Loop Self-Verified Retrieval-Augmented Reasoning for Aerial Vision-Language Navigation

DGX agent

arXiv:2606.28397v1 Announce Type: cross Abstract: Vision-language navigation (VLN) has recently advanced with large language and multimodal models, enabling agents to follow natural-language instructi

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Deterministic Decisions for High-Stakes AI. A Zero-Egress Pipeline with the Deployability of RAG and the Accuracy of Machine Learning

DGX agent

arXiv:2606.29280v1 Announce Type: cross Abstract: We identify intervention bias as a previously unquantified failure mode of zero-shot large-language-model (LLM) educational advisory agents: without t

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

EVAF: A Test-Retest Protocol for Selective Parametric Consolidation

DGX agent

arXiv:2606.29916v1 Announce Type: cross Abstract: Long-running language agents need mechanisms for deciding which experiences should persist after the working context is gone. Retrieval systems can re

model-releasesarxiv-cs-ai
30 Jun 2026
Research

Exploring the Cryptographic Limits of Transformer Networks

DGX agent

arXiv:2606.29389v1 Announce Type: cross Abstract: In recent work it has been shown that colluding AI agents can use steganographic methods to exchange malicious information. Whether a transformer can

researcharxiv-cs-lg
30 Jun 2026
Safety

FutureNav: Unified World-Action Modeling for Vision-and-Language Navigation

DGX agent

arXiv:2606.30367v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) in continuous environments requires an agent to ground instructions in egocentric observations while maintaining sp

safetyarxiv-cs-ro
30 Jun 2026
Safety

Hierarchical Decision Making with Structured Policies: A Principled Design via Inverse Optimization

DGX agent

arXiv:2606.28764v1 Announce Type: new Abstract: Hierarchical decision-making frameworks are pivotal for addressing complex control tasks, enabling agents to decompose intricate problems into manageabl

safetyarxiv-cs-lg
30 Jun 2026
Model Releases

Internal-State Probes Read the Situation, Not the Action: Three Negative Results for Pre-Action Misalignment Monitoring

DGX agent

arXiv:2606.30449v1 Announce Type: new Abstract: Probes on model internals could help monitor agentic systems if they identify harmful text or tool actions before those actions are generated. We ask wh

model-releasesarxiv-cs-lg
30 Jun 2026
Safety

Learned Coordination Conventions in Cooperative MARL: Measuring the Translation Gap Between Theory-Informed Roles and Learned Routing

DGX agent

arXiv:2606.29541v1 Announce Type: new Abstract: Role-semantic assignments provide priors over how heterogeneous agents may coordinate, but cooperative MARL systems instead settle on conventions throug

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Lost in Execution: On the Multilingual Robustness of Tool Calling in Large Language Models

DGX agent

arXiv:2601.05366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly deployed as agents that invoke external tools through structured function calls. While recent wo

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

MirrorCode: AI can rebuild entire programs from behavior alone

DGX agent

arXiv:2606.30182v1 Announce Type: new Abstract: AI models are rapidly improving at autonomous coding, as shown by benchmark progress and one-off demonstrations such as AI implementing a C compiler. Ho

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

DGX agent

arXiv:2606.29265v1 Announce Type: new Abstract: Reasoning large language models (LLMs) have recently made much progress in complex problem-solving, leveraging internal reasoning (or thought) to guide

safetyarxiv-cs-cl
30 Jun 2026
Safety

Modification-Considering Value Learning for Reward Hacking Mitigation in RL

DGX agent

arXiv:2606.28955v1 Announce Type: cross Abstract: Reinforcement learning agents can exploit misspecified reward signals to achieve high apparent returns while failing on the intended objective, a fail

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Optimizing Expert-Designed Crystal Graph Networks for Band-Gap Prediction with an Autonomous LLM Research Loop

DGX agent

arXiv:2606.29717v1 Announce Type: cross Abstract: Predicting a material's properties from its structure is a central, fast-advancing problem in computational materials science. A decade of work has pr

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation

DGX agent

arXiv:2606.29934v1 Announce Type: new Abstract: Image-goal navigation is a key challenge in embodied robotics, where an agent must reach a target specified solely by a goal image. While existing reinf

safetyarxiv-cs-ro
30 Jun 2026
Model Releases

SABER-Math: Automated Benchmark for Information Retrieval Evaluation in Mathematics

DGX agent

arXiv:2606.29894v1 Announce Type: cross Abstract: As agentic AI systems tackle more complex mathematical tasks, they increasingly rely on information retrieval (IR) to search problem databases, theore

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Safety from Honesty in a Disinterested AI Predictor

DGX agent

arXiv:2606.29657v1 Announce Type: new Abstract: As AI systems become more capable, training procedures that optimize for downstream outcomes risk introducing implicit agency: goal-directed behavior th

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

Self-Supervised Theorem Discovery in a Formal Axiomatic System

DGX agent

arXiv:2606.28747v1 Announce Type: new Abstract: Recent artificial intelligence (AI) systems have shown remarkable progress in mathematical reasoning. Many existing approaches, including large language

model-releasesarxiv-cs-ai
30 Jun 2026
Research

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech

DGX agent

arXiv:2606.30543v1 Announce Type: cross Abstract: With the proliferation of speech AI agents, understanding emotional entrainment in conversational interaction has become increasingly important. Emoti

researcharxiv-cs-ai
30 Jun 2026
Model Releases

Translating Natural Language to Strategic Temporal Specifications via LLMs

DGX agent

arXiv:2606.30441v1 Announce Type: cross Abstract: A rigorous formalization of system requirements is a fundamental prerequisite for the verification of Multi-Agent Systems (MAS). However, writing corr

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

X-Mind: Efficient Visual Chain-of-Thought via Predictive World Model for End-to-End Driving

DGX agent

arXiv:2606.28758v1 Announce Type: cross Abstract: Predicting future states is essential for autonomous agents, yet current Vision-Language-Action (VLA) models fundamentally lack this capability, relyi

safetyarxiv-cs-ai
30 Jun 2026
Model Releases

EXPLORE-Bench: Egocentric Scene Prediction with Long-Horizon Reasoning

DGX agent

arXiv:2603.09731v3 Announce Type: replace-cross Abstract: Multimodal large language models (MLLMs) are increasingly considered as a foundation for embodied agents, yet it remains unclear whether they

model-releasesarxiv-cs-ai
29 Jun 2026
Safety

hia-gat: A Heterogeneous Interaction-Aware Graph Attention Network For Frame-Level Traffic Conflict Risk Prediction On Freeways

DGX agent

arXiv:2606.27577v1 Announce Type: cross Abstract: This paper formulates frame-level freeway risk assessment as a multi-agent scene graph-level binary classification problem, where each video or trajec

safetyarxiv-cs-ai
29 Jun 2026
Tutorials

RAE-NWM: Navigation World Model in Dense Visual Representation Space

DGX agent

arXiv:2603.09241v2 Announce Type: replace Abstract: Visual navigation requires agents to reach goals in complex environments through perception and planning. World models address this task by simulati

tutorialsarxiv-cs-cv
29 Jun 2026
Applications

Towards Evaluation of Implicit Software World Models in Coding LLMs

DGX agent

arXiv:2606.27406v1 Announce Type: cross Abstract: Software engineering, whether performed by humans or by AI agents, requires reasoning about how software behaves. We call the internal model that supp

applicationsarxiv-cs-ai
29 Jun 2026
Safety

Towards Value-Constrained Credit Assignment in Fully Delegated AI Cooperatives

DGX agent

arXiv:2606.28217v1 Announce Type: cross Abstract: We propose a framework for reward allocation in fully delegated AI cooperatives where humans are represented by agents that contribute data and partic

safetyarxiv-cs-ai
29 Jun 2026
Local Ai

Understanding Rollout Error in Graph World Models

DGX agent

arXiv:2606.27780v1 Announce Type: new Abstract: World models are often used for planning by rolling learned dynamics forward. Many planning environments, however, are not vectors or images; they are g

local-aiarxiv-cs-ai
29 Jun 2026
Safety

Automating Potential-based Reward Shaping with Vision Language Model Guidance

DGX agent

arXiv:2606.27180v1 Announce Type: cross Abstract: Sparse rewards are inherently challenging for reinforcement learning agents as they lack intermediate feedback to guide exploration and to correctly a

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

Parametric Open Source Games

DGX agent

arXiv:2606.27068v1 Announce Type: cross Abstract: Open-source game theory studies agents whose behavior may depend on one another's decision procedures, but most existing models use discrete or symbol

model-releasesarxiv-cs-ai
26 Jun 2026
Safety

Proposal-Conditioned Latent Diffusion for Closed-Loop Traffic Scenario Generation

DGX agent

arXiv:2606.27123v1 Announce Type: cross Abstract: Closed-loop traffic simulation remains challenging because it must generate interactive multi-agent behaviors that are scene-consistent and controllab

safetyarxiv-cs-cv
26 Jun 2026
Safety

Radical AI Interpretability

DGX agent

arXiv:2606.26523v1 Announce Type: new Abstract: We develop a framework for interpreting AI systems as agents, drawing on the philosophical tradition of radical interpretation and the tools of mechanis

safetyarxiv-cs-ai
26 Jun 2026
Model Releases

SciFig: Towards Automating Editable Figure Generation for Scientific Papers

DGX agent

arXiv:2601.04390v2 Announce Type: replace Abstract: High-quality methodology figures are central to scientific communication, yet they remain difficult and time-consuming to create. Such figures must

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Beyond One-Size-Fits-All: Diagnosis-Driven Online Reinforcement Learning with Offline Priors

DGX agent

arXiv:2606.25527v1 Announce Type: new Abstract: Online reinforcement learning (RL) agents increasingly depend on knowledge acquired offline to achieve practical efficiency. Originally studied in offli

model-releasesarxiv-cs-lg
25 Jun 2026
Model Releases

Improving Zero-Shot Offline RL via Behavioral Task Sampling

DGX agent

arXiv:2604.25496v2 Announce Type: replace Abstract: Offline zero-shot reinforcement learning (RL) aims to learn agents that optimize unseen reward functions without additional environment interaction.

model-releasesarxiv-cs-ai
25 Jun 2026
Safety

SAGE-Nav: Leveraging LLM Planning and Alignment Fusion for Hierarchical Scene Graph-Guided Navigation

DGX agent

arXiv:2606.25497v1 Announce Type: new Abstract: Object-Goal Navigation (ObjNav) requires embodied agents to autonomously locate specified targets using only egocentric visual observations. Existing mo

safetyarxiv-cs-ro
25 Jun 2026
Model Releases

Spatio-Temporal Mixture-of-Modality-Experts Diffusion for Quantitative DCE-MRI Synthesis from Incomplete MR Sequences

DGX agent

arXiv:2606.25535v1 Announce Type: new Abstract: Quantitative maps from dynamic contrast-enhanced MRI (DCE-MRI) are essential for tumor assessment but are often unavailable due to contrast-agent risks

model-releasesarxiv-cs-cv
25 Jun 2026
Model Releases

USS: Unified Spatial-Semantic Prompts for Embodied Visual Tracking with Latent Dynamics Learning

DGX agent

arXiv:2606.25880v1 Announce Type: new Abstract: Embodied Visual Tracking (EVT) requires an agent to continuously follow a specified target while actively moving through dynamic environments. However,

model-releasesarxiv-cs-cv
25 Jun 2026
Safety

Why Multi-Step Tool-Use Reinforcement Learning Collapses and How Supervisory Signals Fix It

DGX agent

arXiv:2606.26027v1 Announce Type: new Abstract: Tool use enables large language models (LLMs) to perform complex tasks, and recent agentic reinforcement learning (RL) methods show promise for enhancin

safetyarxiv-cs-cl
25 Jun 2026
Local Ai

WinDOM: Self-Family Distillation for Small-Model GUI Grounding

DGX agent

arXiv:2606.25964v1 Announce Type: cross Abstract: Small (sim2B) GUI-grounding agents are attractive for on-device deployment, accessibility tooling, and low-cost iteration, but at this scale they face

local-aiarxiv-cs-lg
25 Jun 2026
Model Releases

Learning to Trigger: Reinforcement Learning at the Large Hadron Collider

DGX agent

arXiv:2606.23993v1 Announce Type: cross Abstract: High-throughput scientific facilities such as the Large Hadron Collider depend on real-time event filtering (extit{triggering}) under tight constraint

model-releasesarxiv-cs-ai
24 Jun 2026
Research

The mathbf{P}-Completeness of Inverted Index Traversal: On the Complexity of Evaluating Boolean Query DAGs

DGX agent

arXiv:2601.18747v2 Announce Type: replace-cross Abstract: Modern AI agents increasingly rely on search infrastructure to execute complex, neuro-symbolic reasoning workflows. These workflows often comp

researcharxiv-cs-ai
24 Jun 2026
Research

Beyond the Next Step: Variable-Length Latent World Models for Long-Horizon Planning

DGX agent

arXiv:2606.21775v1 Announce Type: new Abstract: Recently, world models have emerged as a promising paradigm for building intelligent agents by learning predictive models that estimate future environme

researcharxiv-cs-lg
23 Jun 2026
Hardware

Concordia: JIT-Compiled Persistent-Kernel Checkpointing for Fault-Tolerant LLM Inference

DGX agent

arXiv:2606.23521v1 Announce Type: cross Abstract: Long-running LLM agents keep valuable state resident on GPUs: KV caches, request schedulers, communication state, and sometimes online adapters. Losin

hardwarearxiv-cs-lg
23 Jun 2026
Model Releases

Decoupling the Declarative from the Procedural in Vision-Language-Action Models

DGX agent

arXiv:2606.21496v1 Announce Type: cross Abstract: Deploying generalist robotic agents in the real world requires transferable skills. Specifically, a policy trained to clone a behavior from object-spe

model-releasesarxiv-cs-cv
23 Jun 2026
← Previous
1…175176177178179…233
Next →