AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Local Ai

Real-Time Model Checking for Closed-Loop Robot Reactive Planning

DGX agent

arXiv:2508.19186v2 Announce Type: replace-cross Abstract: Reactive obstacle avoidance methods often cause agents to become trapped in local minima, because they can often only reason one step ahead (i

local-aiarxiv-cs-ai
15 Jul 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

The GEST-Engine: From Event Graphs to Synthetic Video. A Full Technical Report

DGX agent

arXiv:2607.12231v1 Announce Type: new Abstract: We present the GEST-Engine, a complete system that goes from natural-language text to fully-annotated multi-actor video. At its core is an explicit worl

agentsarxiv-cs-cv
15 Jul 2026
Safety

When Does Reward Teach State? A Hidden-Automaton Instrument and the Group-Language Boundary

DGX agent

arXiv:2607.11953v1 Announce Type: new Abstract: Does a reinforcement-learning agent that earns high reward represent its task's latent state, or only a reward-correlated shortcut? The question is usua

safetyarxiv-cs-lg
15 Jul 2026
Model Releases

Persuasion Attacks Can Decrease Effectiveness of CoT Monitoring

DGX agent

arXiv:2607.08066v1 Announce Type: new Abstract: Chain-of-thought (CoT) monitoring is a promising safety mechanism for AI agents, based on the premise that visible reasoning traces can surface misalign

model-releasesarxiv-cs-ai
10 Jul 2026
Agents

Akashic: A Low-Overhead LLM Inference Service with MemAttention

DGX agent

arXiv:2607.05708v1 Announce Type: new Abstract: Recent LLM-based agent systems continuously accumulate context across multi-turn interactions, tool invocations, and cross-session workflows. Replaying

agentsarxiv-cs-ai
8 Jul 2026
Agents

CanvasAgent: Enabling Complex Image Creation and Editing via Visual Tool Orchestration

DGX agent

arXiv:2607.05465v1 Announce Type: cross Abstract: Complex image creation and editing often require more than a single generation or editing model. A user request may involve synthesizing images, local

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

EAGOR: Embodied Reasoning in Omni-direction

DGX agent

arXiv:2607.06165v1 Announce Type: new Abstract: Omni-directional (360{eg}) cameras provide embodied agents with a holistic view of their surroundings, making them suited for directional reasoning in t

model-releasesarxiv-cs-ro
8 Jul 2026
Safety

KAT-Coder-V2.5 Technical Report

DGX agent

arXiv:2607.05471v1 Announce Type: cross Abstract: We present KAT-Coder-V2.5, a coding-focused agentic model trained to act autonomously inside real, executable repositories rather than as a single-tur

safetyarxiv-cs-ai
8 Jul 2026
Agents

Unicode TAG-Block Concealment of Tool-Metadata Payloads in the Model Context Protocol: An Approval-View Fidelity Gap Across Three Independent Server Implementations

DGX agent

arXiv:2607.05744v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) is the dominant way coding agents discover and invoke external tools. A server advertises each tool through a tools/l

agentsarxiv-cs-ai
8 Jul 2026
Agents

When Does Tool Use Increase the Expressive Power of Finite-Precision Recurrent Models?

DGX agent

arXiv:2607.06155v1 Announce Type: cross Abstract: Modern sequence models are increasingly deployed as agents that interleave token generation with calls to external tools. We give an exact, architectu

agentsarxiv-cs-cl
8 Jul 2026
Agents

Incentivizing Vision Language Models to Search for Long Video Question Answering

DGX agent

arXiv:2607.02959v1 Announce Type: new Abstract: We introduce VSeek, an agentic framework that transforms long-video question answering (LVQA) from a passive, single-pass perception task into a multi-t

agentsarxiv-cs-cv
7 Jul 2026
Safety

Multi-Turn On-Policy Distillation with Prefix Replay

DGX agent

arXiv:2607.04763v1 Announce Type: cross Abstract: We study on-policy distillation (OPD) for agentic tasks, where an LLM agent interacts with an environment over multiple turns and a student imitates a

safetyarxiv-cs-ai
7 Jul 2026
Agents

Self-Specializing Vision-Language Transmon Chip Calibration in a Physics-Grounded Environment

DGX agent

arXiv:2607.03193v1 Announce Type: cross Abstract: Calibrating a superconducting transmon chip is a sequential decision problem under noise, drift, and a finite budget: an expert must choose experiment

agentsarxiv-cs-ai
7 Jul 2026
Agents

Decentralized Stochastic Subgradient-type Methods with Communication Compression for Nonsmooth Nonconvex Optimization

DGX agent

arXiv:2607.01755v1 Announce Type: cross Abstract: In this paper, we consider the nonsmooth nonconvex decentralized optimization problem, where inter-agent communication is compressed. We propose a gen

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

Distributed Attacks in Persistent-State AI Control

DGX agent

arXiv:2607.02514v1 Announce Type: new Abstract: As AI coding agents become more autonomous, they increasingly ship code iteratively, with the codebase persisting across sessions. This persistence crea

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Lynx: Progressive Speculative Quantization for accelerating KV Transfer in Long-Context Inference

DGX agent

arXiv:2607.01831v1 Announce Type: cross Abstract: Long-context inference is increasingly common in large language model (LLM) serving, driven by retrieval-augmented generation and agentic systems. In

agentsarxiv-cs-lg
3 Jul 2026
Model Releases

Path-level Hindsight Instructions for Semantic Exploration in Vision-Language Navigation

DGX agent

arXiv:2607.01754v1 Announce Type: new Abstract: On-policy exploration is a crucial component for training robust Vision-Language Navigation agents, as it exposes the policy to a broader state distribu

model-releasesarxiv-cs-ai
3 Jul 2026
Agents

Optimal Resource Utilization for Autonomous Laboratory Orchestrators

DGX agent

arXiv:2607.01188v1 Announce Type: new Abstract: In autonomous laboratories, AI agents suggest the next batch of experiments to do. However, planning and executing those tasks taking full advantage of

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

RC-GeoCP: Geometric Consensus for Radar-Camera Collaborative Perception

DGX agent

arXiv:2603.00654v3 Announce Type: replace Abstract: Collaborative perception (CP) improves scene understanding through multi-agent information sharing, yet LiDAR-centric systems remain costly and vuln

model-releasesarxiv-cs-cv
2 Jul 2026
Agents

A Modular Vision-Language-Action Robotics Framework for Indoor Environments

DGX agent

arXiv:2606.31144v1 Announce Type: cross Abstract: This paper presents an integrated system for the CMU Vision-Language-Action (VLA) Challenge, designed to enable an autonomous agent to perform complex

agentsarxiv-cs-ai
1 Jul 2026
Agents

From Similarity to Vulnerability: Key Collision Attack on LLM Semantic Caching

DGX agent

arXiv:2601.23088v2 Announce Type: replace-cross Abstract: Semantic caching has emerged as a pivotal technique for scaling LLM applications, widely adopted by major providers including AWS and Microsof

agentsarxiv-cs-ai
1 Jul 2026
Agents

Plan Right, Then Plan Tight: Symbolic RL for Efficient Embodied Reasoning

DGX agent

arXiv:2606.31260v1 Announce Type: new Abstract: Embodied task planning asks an agent to turn a natural-language instruction into an executable sequence of actions in a physical scene, and is a buildin

agentsarxiv-cs-ro
1 Jul 2026
Agents

The Consistency Dilemma in LLMs: Generator-Evaluator Agreement and Vulnerability to Mistakes

DGX agent

arXiv:2606.30653v1 Announce Type: cross Abstract: Large language models are increasingly deployed in agentic pipelines that depend on the model evaluating its own outputs without external verification

agentsarxiv-cs-ai
1 Jul 2026
Model Releases

Theory of Mind and Persuasion Beyond Conversation: Assessing the Capacity of LLMs to Induce Belief States via Planning and Action

DGX agent

arXiv:2606.31916v1 Announce Type: new Abstract: Theory of Mind (ToM) benchmarks for Large Language Models (LLMs) typically rely on passive question-answering formats, but the deployment of LLMs in inc

model-releasesarxiv-cs-cl
1 Jul 2026
Model Releases

Why Solve It Twice? Hierarchical Accumulation of Skills for Transfer-Efficient ML Engineering

DGX agent

arXiv:2606.30911v1 Announce Type: new Abstract: ML engineering agents waste compute rediscovering known techniques because every competition is a cold start. We present HASTE, a hierarchical multi-age

model-releasesarxiv-cs-ai
1 Jul 2026
Agents

A Cognition-Emotion-Personality Framework for Modeling Human-Like Awareness and Behavior in Emergency Evacuations

DGX agent

arXiv:2606.29212v1 Announce Type: new Abstract: Agent-based evacuation simulations are widely used to study crowd behavior during emergencies, but many models rely on assumptions such as perfect event

agentsarxiv-cs-ai
30 Jun 2026
Agents

DRIVE-Nav: Directional Reasoning, Inspection, and Verification for Efficient Open-Vocabulary Navigation

DGX agent

arXiv:2603.28691v2 Announce Type: replace Abstract: Open-Vocabulary Object Navigation (OVON) requires an embodied agent to locate a language-specified target in unknown environments. Many zero-shot me

agentsarxiv-cs-ro
30 Jun 2026
Agents

Efficient-VLN: A Simple yet Strong Baseline for Efficient Vision-Language Navigation

DGX agent

arXiv:2512.10310v2 Announce Type: replace Abstract: While Multimodal Large Language Models (MLLMs) have demonstrated significant promise in Vision-Language Navigation (VLN), existing agents remain hea

agentsarxiv-cs-cv
30 Jun 2026
Agents

Shell-Supervised Gaussian Splatting for Urban Real-to-Sim Reconstruction

DGX agent

arXiv:2606.30014v1 Announce Type: new Abstract: Real-to-sim reconstruction for embodied AI requires geometry that is useful for collision reasoning, navigation, and agent-environment interaction, not

agentsarxiv-cs-cv
30 Jun 2026
Agents

SICAGE: Speaker-Independent Culture-Aware Gesture Generation using TED4C-L Dataset

DGX agent

arXiv:2606.30001v1 Announce Type: new Abstract: Recent co-speech gesture generation methods often overlook cultural differences, limiting their effectiveness in human-agent interaction. Moreover, cult

agentsarxiv-cs-cv
30 Jun 2026
Agents

The Emergence of Autonomous Penetration Capabilities in Large Language Model-Powered AI Systems

DGX agent

arXiv:2606.13079v2 Announce Type: replace-cross Abstract: Nowadays, the autonomous execution of cyberattacks capable of causing substantial real-world harm is widely regarded as one of the critical re

agentsarxiv-cs-ai
30 Jun 2026
Agents

A Comprehensive Survey on World Models for Embodied AI

DGX agent

arXiv:2510.16732v3 Announce Type: replace Abstract: Embodied AI requires agents that perceive, act, and anticipate how actions reshape future world states. World models serve as internal simulators th

agentsarxiv-cs-cv
29 Jun 2026
Agents

Bridging Talk and Thought: Understanding Dialogue Dynamics Across Collaborative Problem-Solving Contexts

DGX agent

arXiv:2606.27233v1 Announce Type: cross Abstract: We present a conceptual framework for analyzing dialogue in collaborative problem-solving contexts, with an emphasis on the emerging dynamics of human

agentsarxiv-cs-ai
26 Jun 2026
Agents

Learning User Simulators with Turing Rewards

DGX agent

arXiv:2606.19336v2 Announce Type: replace Abstract: Learning to simulate human users in interactive settings could advance the training of agent assistants, evaluation of personalization systems, rese

agentsarxiv-cs-cl
26 Jun 2026
Agents

Reinforcement Learning Enables Autonomous Microrobot Navigation and Intervention in Simulated Blood Capillaries

DGX agent

arXiv:2606.26154v1 Announce Type: cross Abstract: Autonomous microrobots navigating biological vasculature could enable targeted drug delivery and thrombolysis, yet training control policies for reali

agentsarxiv-cs-lg
26 Jun 2026
Agents

Unbiased Canonical Set-Valued Oracles Via Lattice Theory

DGX agent

arXiv:2606.26418v1 Announce Type: new Abstract: A non-agentic 'oracle' AI that estimates probabilities of future events faces a self-reference problem: once its answer is learned and acted upon, it ca

agentsarxiv-cs-ai
26 Jun 2026
Agents

FactorLibrary: From Polynomials to Circuits via Recursive Subgoals

DGX agent

arXiv:2606.25394v1 Announce Type: new Abstract: Finding minimal arithmetic circuits for polynomials over finite fields is a combinatorially hard problem central to algebraic complexity theory. We form

agentsarxiv-cs-lg
25 Jun 2026
Agents

Evolving Programmatic Skill Networks

DGX agent

arXiv:2601.03509v2 Announce Type: replace Abstract: We study continual skill acquisition in open-ended embodied environments where an agent must construct, refine, and reuse an expanding library of ex

agentsarxiv-cs-ai
24 Jun 2026
Agents

Forget Without Compromise: Nexus Sampling for Streaming KV-Cache Eviction Under Fixed Budgets

DGX agent

arXiv:2606.23961v1 Announce Type: new Abstract: Long-context and agentic LLM workloads push the KV cache past any fixed memory budget, forcing the inference stack to permanently evict tokens at every

agentsarxiv-cs-lg
24 Jun 2026
Safety

When AI Meets Finance (StockAgent): Large Language Model-based Stock Trading in Simulated Real-world Environments

DGX agent

arXiv:2407.18957v5 Announce Type: replace-cross Abstract: Can AI Agents simulate real-world trading environments to investigate the impact of external factors on stock trading activities (e.g., macroe

safetyarxiv-cs-ai
24 Jun 2026
Local Ai

A Gated Graph Neural Network Approach to Fast-Convergent Dynamic Average Estimation

DGX agent

arXiv:2606.20955v1 Announce Type: new Abstract: Dynamic average estimation is a critical problem in multi-agent systems, enabling agents to collaboratively estimate time-varying signals using only loc

local-aiarxiv-cs-lg
23 Jun 2026
Agents

DUET: Decentralized Bilevel Optimization without Lower-Level Strong Convexity

DGX agent

arXiv:2606.21153v1 Announce Type: cross Abstract: Decentralized bilevel optimization (DBO) provides a powerful framework for multi-agent systems to solve local bilevel tasks in a decentralized fashion

agentsarxiv-cs-lg
23 Jun 2026
Agents

Equilibrium with Internal Transfers

DGX agent

arXiv:2606.20960v1 Announce Type: cross Abstract: Nash equilibrium (NE) arises from selfish utility maximization, yet its social welfare can be arbitrarily far from optimal. Moreover, computing an NE

agentsarxiv-cs-lg
23 Jun 2026
Safety

Inverting the Bellman Equation: From Q-Values to World Models

DGX agent

arXiv:2606.21173v1 Announce Type: new Abstract: Model-based and model-free reinforcement learning are traditionally viewed as separate paradigms: instead of learning a model of the transition kernel P

safetyarxiv-cs-lg
23 Jun 2026
Agents

Negative Knowledge as Failure-aware Shared Memory for AutoResearch

DGX agent

arXiv:2606.21024v1 Announce Type: cross Abstract: AI-assisted research systems generate many failed attempts, but those failures rarely become a durable, shared knowledge asset. We propose a negative

agentsarxiv-cs-lg
23 Jun 2026
Agents

Towards Responsibly Non-Compliant Machines

DGX agent

arXiv:2606.12147v1 Announce Type: new Abstract: We consider the problem of engineering autonomous intelligent agents that are capable to responsibly not comply with user requests. We argue that machin

agentsarxiv-cs-ai
11 Jun 2026
Agents

When More Documents Hurt RAG: Mitigating Vector Search Dilution with Domain-Scoped, Model-Agnostic Retrieval

DGX agent

arXiv:2606.11350v1 Announce Type: new Abstract: Retrieval-augmented generation degrades when scaled to large, heterogeneous document collections, where dense similarity loses discriminative power, and

agentsarxiv-cs-cl
11 Jun 2026
Agents

EstRTL: Functional Estimation Guided RTL Code Generation

DGX agent

arXiv:2606.09867v1 Announce Type: cross Abstract: Optimizing register transfer level (RTL) code is of vital importance in hardware design. Large language models (LLMs) provide new methods for the auto

agentsarxiv-cs-ai
10 Jun 2026
← Previous
1…128129130131132…236
Next →