AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

Expected Free Energy-based Planning as Variational Inference

DGX agent

arXiv:2606.20658v1 Announce Type: cross Abstract: Planning under uncertainty requires agents to balance goal achievement with information gathering. Active inference addresses this through the Expecte

safetyarxiv-cs-lg
23 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Fair Transit Stop Placement: A Clustering Perspective and Beyond

DGX agent

arXiv:2602.06776v2 Announce Type: replace-cross Abstract: We study the transit stop placement (TrSP) problem in general metric spaces, where agents travel between source-destination pairs and may eith

safetyarxiv-cs-lg
23 Jun 2026
Applications

FlowDec: Temporal Conditional Flow Decorruptor for Robust Continuous Vision-Language Navigation

DGX agent

arXiv:2606.22424v1 Announce Type: new Abstract: Vision-and-Language Navigation in Continuous Environments (VLN-CE) requires agents to follow natural-language instructions in unseen scenes. While Large

applicationsarxiv-cs-cv
23 Jun 2026
Model Releases

GRAG: Generic Response-Augmented Generation Framework for Personalized Conversational Systems

DGX agent

arXiv:2606.21097v1 Announce Type: cross Abstract: Deploying highly capable personalized conversational agents in resource-constrained or privacy-sensitive environments remains a significant challenge.

model-releasesarxiv-cs-lg
23 Jun 2026
Safety

Horizon Adaptive Offline Policy Learning via Value Stitching

DGX agent

arXiv:2606.21136v1 Announce Type: new Abstract: Learning accurate value functions plays a decisive role for reinforcement learning (RL) agents to solve long-horizon, complex tasks. Conventional tempor

safetyarxiv-cs-lg
23 Jun 2026
Safety

Inductive Generalization for Robotic Manipulation

DGX agent

arXiv:2606.20999v1 Announce Type: cross Abstract: Understanding the generalization capabilities of visuomotor policies is essential in the development of capable robotic agents. Generalizable models l

safetyarxiv-cs-lg
23 Jun 2026
Model Releases

IOI: Decoupling Kinematics and Physics for Interactive World Models

DGX agent

arXiv:2606.23296v1 Announce Type: new Abstract: Developing generalist embodied agents requires interactive environments providing visually realistic feedback and accurate action-conditioned dynamics.

model-releasesarxiv-cs-ro
23 Jun 2026
Safety

Learned Controllers for Agile Quadrotors in Pursuit-Evasion Games

DGX agent

arXiv:2506.02849v3 Announce Type: replace-cross Abstract: In this letter we study 1v1 quadrotor pursuit-evasion, where a pursuer and an evader are trained via reinforcement learning (RL) by competing

safetyarxiv-cs-lg
23 Jun 2026
Hardware

SCENIC: Semantic-Conditioned Edge-Aware Neural Framework for Structured IoT Command Generation

DGX agent

arXiv:2606.22296v1 Announce Type: new Abstract: Edge Internet of Things (IoT) agents are often constrained by memory capacity, privacy requirements, communication latency, and recurring inference cost

hardwarearxiv-cs-lg
23 Jun 2026
Research

Stealthy World Model Manipulation via Data Poisoning

DGX agent

arXiv:2606.18697v2 Announce Type: replace Abstract: Model-based learning agents use learned world models to predict future states, plan actions, and adapt to new environments. However, the process of

researcharxiv-cs-lg
23 Jun 2026
Safety

Architecture-Aware Reinforcement Learning Makes Sliding-Window Attention Competitive in Math Reasoning

DGX agent

arXiv:2606.11634v1 Announce Type: new Abstract: The rapid progress of reasoning and agentic large language models (LLMs) has increased the demand for long-context inference, but self-attention (SA) sc

safetyarxiv-cs-ai
11 Jun 2026
Research

Context-Driven Incremental Compression for Multi-Turn Dialogue Generation

DGX agent

arXiv:2606.12411v1 Announce Type: new Abstract: Modern conversational agents condition on an ever-growing dialogue history at each turn, incurring redundant attention and encoding costs that grow with

researcharxiv-cs-cl
11 Jun 2026
Model Releases

FronTalk: Benchmarking Front-End Development as Conversational Code Generation with Multi-Modal Feedback

DGX agent

arXiv:2601.04203v2 Announce Type: replace Abstract: We present FronTalk, a benchmark for front-end code generation that pioneers the study of a unique interaction dynamic: conversational code generati

model-releasesarxiv-cs-cl
11 Jun 2026
Model Releases

How Auxiliary Reasoning Unleashes GUI Grounding in VLMs

DGX agent

arXiv:2509.11548v2 Announce Type: replace Abstract: Graphical user interface (GUI) grounding is a fundamental task for building GUI agents. However, general vision-language models (VLMs) struggle with

model-releasesarxiv-cs-cv
11 Jun 2026
Safety

Learning to Inject: Automated Prompt Injection via Reinforcement Learning

DGX agent

arXiv:2602.05746v2 Announce Type: replace-cross Abstract: Prompt injection is a critical vulnerability in LLM agents, yet the strongest methods still rely on human red-teamers and hand-crafted prompts

safetyarxiv-cs-ai
11 Jun 2026
Model Releases

Mind the Perspective: Let's Reason Recursively for Theory of Mind

DGX agent

arXiv:2606.11724v1 Announce Type: new Abstract: Theory of Mind (ToM) reasoning requires inferring agents' beliefs from partial and asymmetric observations, which remains an open challenge for LLMs. Ex

model-releasesarxiv-cs-ai
11 Jun 2026
Local Ai

PIGEON: VLM-Driven Object Navigation via Points of Interest Selection

DGX agent

arXiv:2511.13207v2 Announce Type: replace-cross Abstract: Object navigation in unseen indoor environments requires agents to perform semantic search under partial observability. Vision-language models

local-aiarxiv-cs-cv
11 Jun 2026
Model Releases

TouchThinker: Scaling Tactile Commonsense Reasoning to the Open World with Large-scale Data and Action-aware Representation

DGX agent

arXiv:2606.11637v1 Announce Type: new Abstract: Touch is a key modality for embodied agents to understand the physical world. Although recent work has incorporated tactile signals into language system

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

A Survey of Robotic Navigation and Manipulation with Physics Simulators in the Era of Embodied AI

DGX agent

arXiv:2505.01458v2 Announce Type: replace-cross Abstract: Navigation and manipulation are core capabilities in Embodied AI, but training agents to perform them directly in the real world is costly, ti

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

CoCoSI: Collaborative Cognitive Map Construction for Spatial Intelligence

DGX agent

arXiv:2606.10401v1 Announce Type: new Abstract: Spatial intelligence is a key frontier for multimodal large language models (MLLMs), enabling them to reason about the physical world from visual experi

model-releasesarxiv-cs-cv
10 Jun 2026
Model Releases

PreAct-Bench: Benchmarking Predictive Monitoring in LLMs

DGX agent

arXiv:2606.09890v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents capable of executing multi-step action trajectories toward a given objecti

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

ReflectiChain: Epistemic Grounding in LLM-Driven World Models for Supply Chain Resilience

DGX agent

arXiv:2606.10359v1 Announce Type: new Abstract: AI agents in supply chains face a fundamental epistemic gap: large language models (LLMs) interpret policies but lack physical grounding, while reinforc

model-releasesarxiv-cs-ai
10 Jun 2026
Safety

Rethinking Embodied Navigation via Relational Inductive Bias

DGX agent

arXiv:2606.10348v1 Announce Type: new Abstract: Object navigation requires an agent to locate a target in an unknown environment through visual observations. Existing methods typically rely on open-vo

safetyarxiv-cs-ro
10 Jun 2026
Model Releases

RoboGPT-R1: Enhancing Robot Task Planning with Reinforcement Learning

DGX agent

arXiv:2510.14828v3 Announce Type: replace Abstract: Improving the reasoning capabilities of embodied agents is crucial for robots to complete complex human instructions in long-view manipulation tasks

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

SHAPO: Sharpness-Aware Policy Optimization for Safe Exploration

DGX agent

arXiv:2606.10228v1 Announce Type: cross Abstract: Safe exploration is a prerequisite for deploying reinforcement learning (RL) agents in safety-critical domains. In this paper, we approach safe explor

model-releasesarxiv-cs-ai
10 Jun 2026
Research

A Geometric Theory of Cognition for Machine Intelligence

DGX agent

arXiv:2512.12225v3 Announce Type: replace Abstract: Developing artificial agents that unify representation, memory, adaptation, and prediction remains a fundamental challenge in artificial intelligenc

researcharxiv-cs-ai
9 Jun 2026
Research

DIJIT: A Robotic Head for an Active Observer

DGX agent

arXiv:2512.07998v2 Announce Type: replace-cross Abstract: We present DIJIT, a novel binocular robotic head expressly designed for mobile agents that behave as active observers. DIJIT's unique breadth

researcharxiv-cs-cv
9 Jun 2026
Model Releases

Language-based Trial and Error Falls Behind in the Era of Experience

DGX agent

arXiv:2601.21754v3 Announce Type: replace Abstract: While Large Language Models (LLMs) excel in language-based agentic tasks, their applicability to unseen, nonlinguistic environments (e.g., symbolic

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

LargeMonitor: Monitoring Online Task-Free Continual Learning via Large Pretrained Models

DGX agent

arXiv:2606.09430v1 Announce Type: cross Abstract: Online task-free continual learning (TFCL) requires intelligent agents to sequentially accumulate knowledge from an unbounded, non-stationary data str

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

LLM Inference at the Edge: Mobile, NPU, and GPU Performance Efficiency Trade-offs Under Sustained Load

DGX agent

arXiv:2603.23640v2 Announce Type: replace-cross Abstract: Deploying large language models on-device for always-on personal agents demands sustained inference from hardware tightly constrained in power

model-releasesarxiv-cs-lg
9 Jun 2026
Hardware

MuJoCo-Drones-Gym: A GPU-Accelerated Multi-Drone Simulator for Control and Reinforcement Learning

DGX agent

arXiv:2606.08039v1 Announce Type: new Abstract: Robotic simulators are a cornerstone of modern research in aerial robotics, serving both as a vehicle for the development of new control algorithms and

hardwarearxiv-cs-ro
9 Jun 2026
Hardware

Towards Automated Kernel Generation in the Era of LLMs

DGX agent

arXiv:2601.15727v3 Announce Type: replace Abstract: The performance of modern AI systems is fundamentally constrained by the quality of their underlying GPU kernels, which translate high-level algorit

hardwarearxiv-cs-lg
9 Jun 2026
Model Releases

Where Instruction Hierarchy Breaks: Diagnosing and Repairing Failures in Reasoning Language Models

DGX agent

arXiv:2606.07808v1 Announce Type: new Abstract: Reasoning language models deployed in agentic workflows must follow an instruction hierarchy: when instructions from different sources conflict, the mod

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

WhiFlash: Accelerating Speculative Decoding with Token-Level Cross-Paradigm Routing

DGX agent

arXiv:2606.07710v1 Announce Type: cross Abstract: The autoregressive nature of large language models (LLMs) remains a significant bottleneck for inference, particularly in complex agentic workloads. W

safetyarxiv-cs-ai
9 Jun 2026
Research

Accelerated Decentralized Stochastic Gradient Descent for Strongly Convex Optimization

DGX agent

arXiv:2606.07496v1 Announce Type: new Abstract: Decentralized stochastic optimization is a fundamental paradigm for large-scale learning over networks, where agents communicate only with their neighbo

researcharxiv-cs-lg
8 Jun 2026
Model Releases

Beyond Waypoints: A Trajectory-Centric Waypointing Paradigm for Vision-Language Navigation

DGX agent

arXiv:2606.07244v1 Announce Type: cross Abstract: Vision-Language Navigation in Continuous Environments (VLN-CE) requires agents to follow natural-language instructions while navigating in real-world-

model-releasesarxiv-cs-ai
8 Jun 2026
Tutorials

CAPE: Contrastive Action-conditioned Parallel Encoding for Embodied Planning

DGX agent

arXiv:2606.07304v1 Announce Type: new Abstract: Embodied agents need to predict the future consequences of candidate actions in order to plan effectively before execution. Existing visual dynamics mod

tutorialsarxiv-cs-ro
8 Jun 2026
Local Ai

Learning Explicit Behavioral Models with Adaptive Questions and World-Model Probes

DGX agent

arXiv:2606.07127v1 Announce Type: new Abstract: Interactive agents trained only against task return can achieve high scores while failing to represent the mechanisms that make their actions succeed. T

local-aiarxiv-cs-lg
8 Jun 2026
Model Releases

NTILC: Neural Tool Invocation via Learned Compression

DGX agent

arXiv:2606.06566v1 Announce Type: cross Abstract: Agentic tool-calling language models depend on large registries of callable APIs, functions, and local actions. Placing full tool specifications direc

model-releasesarxiv-cs-ai
8 Jun 2026
Model Releases

RECAP: Regression Evaluation for Continual Adaptation of Prompts

DGX agent

arXiv:2606.06698v1 Announce Type: cross Abstract: Production agentic systems routinely face evolving constraints and must comply from the very next interaction. Scenarios like a tool-call notification

model-releasesarxiv-cs-cl
8 Jun 2026
Safety

Robust Driving Control for Autonomous Vehicles: An Intelligent General-sum Constrained Adversarial Reinforcement Learning Approach

DGX agent

arXiv:2510.09041v3 Announce Type: replace-cross Abstract: Deep reinforcement learning (DRL) has demonstrated remarkable success in developing autonomous driving policies. However, its vulnerability to

safetyarxiv-cs-ai
8 Jun 2026
Model Releases

Brick-Composer: Using MLLMs for Assembly with Diverse Bricks

DGX agent

arXiv:2606.05445v1 Announce Type: new Abstract: We dream of AI agents that can read arbitrary designs and construct real-world objects from reusable building blocks. As a first step toward this vision

model-releasesarxiv-cs-ai
6 Jun 2026
Model Releases

DragOn: A Benchmark and Dataset for Drag-Based GUI Interactions

DGX agent

arXiv:2606.06322v1 Announce Type: new Abstract: GUI agents - vision-based models that control desktops, web browsers, and mobile devices through graphical user interfaces - promise to automate a wide

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

GIPO: Gaussian Importance Sampling Policy Optimization

DGX agent

arXiv:2603.03955v2 Announce Type: replace-cross Abstract: Post-training with reinforcement learning (RL) has recently shown strong promise for advancing multimodal agents beyond supervised imitation.

safetyarxiv-cs-ai
6 Jun 2026
Model Releases

Goedel-Architect: Streamlining Formal Theorem Proving with Blueprint Generation and Refinement

DGX agent

arXiv:2606.06468v1 Announce Type: new Abstract: We introduce Goedel-Architect, an agentic framework for formal theorem proving in Lean 4 centered on blueprint generation and refinement. A blueprint is

model-releasesarxiv-cs-ai
6 Jun 2026
Safety

Flow-based Policy Adaptation without Policy Updates

DGX agent

arXiv:2606.06461v1 Announce Type: new Abstract: Leveraging prior knowledge from pretrained policies, foundation models, or human operators offers an efficient alternative to learning robot skills from

safetyarxiv-cs-ro
5 Jun 2026
Model Releases

KV-Control: Parameter-Efficient K/V Injection for Trajectory-Controlled Text-to-Motion

DGX agent

arXiv:2606.05624v1 Announce Type: new Abstract: Text-conditioned 3D human motion models now synthesize plausible motions from prompts, but practical animation and embodied-agent workflows rarely stop

model-releasesarxiv-cs-cv
5 Jun 2026
Local Ai

LeanMarathon: Toward Reliable AI Co-Mathematicians through Long-Horizon Lean Autoformalization

DGX agent

arXiv:2606.05400v1 Announce Type: cross Abstract: Long-horizon autoformalization of research mathematics fails not only at hard lemmas, but at scale: statements drift, dependencies tangle, context dec

local-aiarxiv-cs-cl
5 Jun 2026
← Previous
1…176177178179180…233
Next →