AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “arxiv-cs-ro”

GridTimelineEvolution
3,443 results
4 Aug 2026

AffordTrajDP: Dynamic Affordance-Guided Visuomotor Policy Learning for Robotic Manipulation

SafetyDGX agent

arXiv:2608.01603v1 Announce Type: new Abstract: Affordance-guided imitation learning has shown impressive performance in robotic manipulation tasks by compressing visual perception into task-specific

Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks

SafetyDGX agent

arXiv:2602.12244v2 Announce Type: replace Abstract: Open world language conditioned task planning is crucial for robots operating in large-scale household environments. While many recent works attempt

Assistant Placement Aria: A Benchmark for Egocentric Placement Assistance

Model ReleasesDGX agent

arXiv:2608.00652v1 Announce Type: new Abstract: Human assistance in robotics spans around several tasks such as navigation, object manipulation, and placement, where a key challenge is selecting targe


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

bFaaaP: An Inclusive, Head-Angle Piano-Pedal Interaction that Quantitatively Reproduces a Pianist's Intended Pedalling -- Foot-Free, for Acoustic and Electronic Pianos

Local AiDGX agent

arXiv:2608.00633v1 Announce Type: cross Abstract: Expressive piano performance depends on the sustain (damper) pedal, operated by foot, excluding players who cannot readily use their feet: wheelchair

Bicycle Acrobatics with Reinforcement Learning

AgentsDGX agent

arXiv:2608.00880v1 Announce Type: new Abstract: Bicycle robots are fast and energy efficient, but their simple mechanical design and their underactuated and non-holonomic dynamics make highly agile ma

Bridging the Sim-to-Real Gap in Parallel-Link Leg Mechanisms via Simulator-Side Dynamics Normalization

SafetyDGX agent

arXiv:2608.01697v1 Announce Type: new Abstract: This paper addresses the sim-to-real gap in dynamics arising when a parallel-link mechanism is represented by a serial-tree surrogate in simulation. Con

CAAT: Contact-Aware Attention Scaling and Tactile Masking for Data-Efficient Contact-Rich Manipulation

SafetyDGX agent

arXiv:2608.01102v1 Announce Type: new Abstract: In contact-rich manipulation, visual observations primarily guide motion in free space, whereas tactile observations become particularly informative dur

Certifying Plans under Model Mismatch: A Trilemma for Reachability from Scarce Data

Model ReleasesDGX agent

arXiv:2608.02453v1 Announce Type: new Abstract: Sim-to-real policies are designed under nominal dynamics, but target-system trials may yield only a few isolated one-step transitions. We study pre-exec

ChainVLA: Chaining Vision-Language-Action Queries through a Unified Execution State for Long-Horizon Manipulation

Model ReleasesDGX agent

arXiv:2608.02326v1 Announce Type: new Abstract: Humans perform long-horizon manipulation by retaining knowledge of what earlier actions have established while continuously adapting the motion underway

CoLI: A Reproducible Platform for Continuum Robot Learning via Monolithic 3D Printing and Isomorphic Teleoperation

SafetyDGX agent

arXiv:2606.20389v2 Announce Type: replace Abstract: Continuum robots offer strong potential for manipulation tasks due to their high degrees of freedom, compliant structures, and operational safety. H

Complete Motion Planning using Workspace-Fibered Decomposition for nR-Planar Manipulator

ResearchDGX agent

arXiv:2608.01172v1 Announce Type: new Abstract: We propose a workspace-fibered decomposition framework for motion planning in nR planar redundant manipulators operating in cluttered environments. Rath

Compliant Sphere Lattice Contact: Distributed Contact Modeling for Sphere-Based Robot Representations

ResearchDGX agent

arXiv:2608.00263v1 Announce Type: new Abstract: Contact planning in robotics requires models that are both computationally efficient and physically accurate. Sphere-based robot representations satisfy

CoNav-UAV: Cooperative Dual-Altitude Aerial Navigation via Stackelberg Learning

Model ReleasesDGX agent

arXiv:2608.01802v1 Announce Type: cross Abstract: Target-oriented vision-and-language navigation (VLN) on aerial platforms is attracting growing attention for missions such as disaster rescue, infrast

Demystifying When and Why VLAs Fail in Contact-Rich Tasks and How to Fix Them

SafetyDGX agent

arXiv:2608.01402v1 Announce Type: new Abstract: We address the problem of understanding when and why Vision-Language-Action models struggle with contact-rich manipulation tasks that require precise ph

Deployment-Ready UWB Localization for Industrial Ground Robots with Automatic Anchor Calibration and Terrain-Aware Fusion

Model ReleasesDGX agent

arXiv:2607.15807v2 Announce Type: replace Abstract: Ultra-Wideband (UWB) ranging has become a viable option for industrial Autonomous Mobile Robot (AMR) localization due to improved accuracy and low c

Developing Combined Manipulation and Locomotion Skills with Interaction Representation and Skill Composition

SafetyDGX agent

arXiv:2608.00208v1 Announce Type: new Abstract: This paper addresses how to enable a humanoid robot to learn motion policies based on developmental principles and combine policies to create more sophi

Diffusion-Based Body Schema Learning Enabling Abnormal-State Adaptation in Musculoskeletal Robots

TutorialsDGX agent

arXiv:2608.01029v1 Announce Type: new Abstract: Musculoskeletal robots require an internal body schema that remains consistent under a wide range of physical state changes, including abnormalities suc

Disentangling Visuo-Tactile Foresight: Oracle-Guided Interface Discovery for World Action Models

Model ReleasesDGX agent

arXiv:2608.00547v1 Announce Type: new Abstract: Contact-rich manipulation remains challenging because successful control depends on physical interaction cues that are often weakly observable from visi

DPL: Depth-only Perceptive Humanoid Locomotion via Realistic Depth Synthesis and Cross-Attention Terrain Reconstruction

SafetyDGX agent

arXiv:2510.07152v3 Announce Type: replace Abstract: Recent advancements in legged robot perceptive locomotion have shown promising progress. However, terrain-aware humanoid locomotion remains largely

DreamTrajectory: Trajectory-Guided Action Generation with World Model Alignment for Mobile Manipulation

SafetyDGX agent

arXiv:2608.01381v1 Announce Type: new Abstract: Mobile manipulation requires a robot to coordinate base and arm motion under continuously changing viewpoints and contact conditions, within an action s

Dynamic UAV-based search operations using probabilistic diffusion modeling of Man Overboard incident victims

ResearchDGX agent

arXiv:2608.02093v1 Announce Type: new Abstract: More than 70% of the people that fell overboard cruise ships in the period 2010-2019 lost their lives. This paper presents a strategy for reliably predi

DynamicWAM: Dual-Path Motion Conditioning for World-Action Models in Dynamic Manipulation

ApplicationsDGX agent

arXiv:2608.00793v1 Announce Type: new Abstract: Dynamic manipulation requires robots to infer target motion and respond promptly, yet existing World-Action Models (WAMs) typically condition only on th

Ego2Robot: Scalable Robot Data Synthesis from Egocentric Human Data

ResearchDGX agent

arXiv:2608.02580v1 Announce Type: new Abstract: Learning generalizable robot manipulation policies requires large-scale and diverse demonstration data. Egocentric human manipulation videos offer rich

Embodied Passive Aeroacoustic Perception Enables Relative Sensing and Pursuit Between Aerial Robots

ResearchDGX agent

arXiv:2608.00401v1 Announce Type: new Abstract: Aerial robots generate structured aeroacoustic fields during flight, yet these signals have been underexplored as a source of onboard relative perceptio

EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation

SafetyDGX agent

arXiv:2608.01221v1 Announce Type: new Abstract: Autonomous endoscopic navigation can reduce clinicians' operational burden, yet robust control remains challenging due to tissue deformation, transient

Environmental resilience via morphological diversity within machines

ResearchDGX agent

arXiv:2608.02395v1 Announce Type: new Abstract: Organisms contain diverse, sensorimotor parts across size scales and rapidly adapt to new environments, while machines contain only inert materials at s

First Deployable Dynamic-CoM: A Unified Policy and Method-Agnostic Benchmark for Humanoid Single-Leg Balance

Model ReleasesDGX agent

arXiv:2608.00500v1 Announce Type: new Abstract: Unified humanoid policies handle agile whole-body motion, yet stumble on a simple demand: staying balanced on one leg. On our single-leg-balance benchma

FlowPilot: Real-Time World-Action Modeling for Agile UAV Navigation

Local AiDGX agent

arXiv:2608.00635v1 Announce Type: new Abstract: We present FlowPilot, a compact world-action model for real-time onboard UAV navigation from depth. Unlike map-then-optimize pipelines that require loca

FRA-NBV: A Fast and Reflectivity-Aware Next-Best-View Strategy

AgentsDGX agent

arXiv:2608.01950v1 Announce Type: new Abstract: Autonomous 3D reconstruction with depth sensors is strongly affected by reflective surfaces, which cause missing or unreliable measurements and reduce t

FreqNav: Stage-Wise Frequency Routing for Object-Oriented Aerial Vision-Language Navigation

Local AiDGX agent

arXiv:2608.00970v1 Announce Type: new Abstract: Object-oriented aerial vision-and-language navigation (VLN) requires searching for a described target and landing on it precisely, under long-horizon an

From Failures to Supervision: DynamicEnvPlan for Robust Long-Horizon Embodied Planning

SafetyDGX agent

arXiv:2608.00613v1 Announce Type: new Abstract: Physical-world interaction is inherently dynamic, as environments can evolve during execution, requiring agents to adapt their plans under non-stationar

GeminiPainter's sequence-formed pipeline comprised of perception, cognition, planning, and action stages

Model ReleasesDGX agent

arXiv:2608.00829v1 Announce Type: new Abstract: We present an autonomous robotic portrait-generation system combining real-time face detection, AI-based sketch generation, and robotic drawing. The sys

Good Weights: Proactive, Adaptive Dead Reckoning Fusion for Continuous and Robust Visual SLAM

ApplicationsDGX agent

arXiv:2509.22910v2 Announce Type: replace Abstract: Given that Visual SLAM relies on appearance cues for localization and scene understanding, texture-less or visually degraded environments (e.g., pla

Grasp Execution Without a Planner: Configuration-Space Grasp Distance Fields with Certified Safety & Guaranteed Quality

Model ReleasesDGX agent

arXiv:2608.00600v1 Announce Type: new Abstract: Standard multifingered grasp execution architectures plan a collision-free trajectory to a selected grasp pose and track it with a feedback law. Executi

Grounded Semantic Re-Binding for Robust Instruction Generalization in Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2608.02497v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models excel in robotic manipulation but suffer catastrophic performance drops when canonical instructions are simply parap

Grounded Vision-Language Interpreter for Long-Horizon Bimanual Task and Motion Planning

SafetyDGX agent

arXiv:2506.03270v3 Announce Type: replace Abstract: While recent advances in vision-language models have accelerated language-guided robot planning, their black-box nature lacks the safety guarantees

HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing

SafetyDGX agent

arXiv:2603.15257v2 Announce Type: replace Abstract: Tactile sensing is a crucial capability for Vision-Language-Action (VLA) architectures, as it enables dexterous and safe manipulation in contact-ric

Hybrid Attention Estimation Pipeline for Adaptive HRI Using an Expressive Robotic Head

ApplicationsDGX agent

arXiv:2608.00284v1 Announce Type: new Abstract: This paper presents an applied case study on hybrid visual attention estimation for human-robot interaction using an expressive robotic head based on th

Hybrid Impedance-Admittance Control with Multi-Link Aerial Robot for Contact-Rich Surface Sliding Task

ResearchDGX agent

arXiv:2608.01800v1 Announce Type: new Abstract: Multi-link aerial robots can actively deform their articulated structures during flight, giving them strong potential for aerial manipulation. However,

Interaction Dynamics MPC for Knee Rehabilitation Exoskeletons: A Closed-Loop SEA Outer-Loop Study

SafetyDGX agent

arXiv:2606.13485v3 Announce Type: replace-cross Abstract: Safe rehabilitation is an interaction-dynamics problem: the controller must regulate a prescribed motion while absorbing involuntary spasm, vo

KING: Embodiment-Aware Kinematic Graph Neural Network for Unified Motion Representation of Legged and Wheeled Robots

ResearchDGX agent

arXiv:2608.01015v1 Announce Type: new Abstract: Kinematic models provide reliable motion constraints for odometry estimation in featureless environments, where exteroceptive sensing degrades and IMU i

Latency-Tolerant Cloud-Edge Collaborative Vision-Language-Action Models via Emergent Representational Specialization

Model ReleasesDGX agent

arXiv:2608.00569v1 Announce Type: new Abstract: Deploying billion-parameter Vision-Language-Action (VLA) policies on mobile robots creates a systems conflict: semantic reasoning benefits from cloud GP

Learning-Based Motion Planning for Dynamic Environments: From Foundational Algorithms to Emerging Paradigms

SafetyDGX agent

arXiv:2608.00625v1 Announce Type: new Abstract: Motion planning in dynamic environments is a fundamental problem in robotics, aiming to generate safe and efficient paths, trajectories, or control acti

Learning Panorama-Aware VLA for Mobile Manipulation with Whole-Body Teleoperation

SafetyDGX agent

arXiv:2608.02257v1 Announce Type: new Abstract: Mobile manipulation is a key capability for embodied intelligence, enabling robots to accomplish complex multi-stage tasks in open-world environments. H

Learning Smooth SE(3) Trajectories under Left-Invariant Riemannian Metrics

TutorialsDGX agent

arXiv:2608.01562v1 Announce Type: new Abstract: Optimal trajectory generation for rigid-body motions on Lie groups can be formulated as a variational problem that minimizes energy functionals defined

Learning to Predict Contact Force Distributions from Vision Leveraging Object Geometry Priors

ApplicationsDGX agent

arXiv:2608.00464v1 Announce Type: new Abstract: Based on vision and prior experience, humans can make rough physical predictions and adjust their manipulation strategies. This paper aims to endow robo

Local-Canonicalization Equivariant Graph Neural Networks for Sample-Efficient and Generalizable Swarm Robot Control

SafetyDGX agent

arXiv:2509.14431v2 Announce Type: replace Abstract: Multi-agent reinforcement learning (MARL) policies for swarm control often learn inefficiently and generalize poorly across coordinate frames, team

Localization in Spatiotemporal Fields via Environmental PDEs

SafetyDGX agent

arXiv:2608.00272v1 Announce Type: new Abstract: This paper proposes a localization framework that uses spatiotemporal fields governed by partial differential equations (PDEs) as localization signature

Look Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models

Model ReleasesDGX agent

arXiv:2608.02197v1 Announce Type: new Abstract: Visual representations of VLA models remain unreliable for spatially precise robotic manipulation. We uncover that vision encoders in VLAs also exhibit

LooperMuscle: Fast and Stable Learning of Humanoid Whole-Body Tracking via Structured Mixture-of-Experts

Model ReleasesDGX agent

arXiv:2608.00820v1 Announce Type: new Abstract: FastSAC-style methods significantly reduce humanoid motion training time but often suffer from notable performance degradation compared with PPO in whol

MANGO-Grasp: Mahalanobis Fields over Geometry-Oriented 3D Gaussians for Cross-Embodiment Dexterous Grasping

Local AiDGX agent

arXiv:2608.02014v1 Announce Type: new Abstract: Cross-embodiment dexterous grasping aims to synthesize stable grasps across heterogeneous multi-fingered hands with little or no embodiment-specific tun

MemoAct: Atkinson-Shiffrin-Inspired Hierarchical Memory-Augmented Policy for Robotic Manipulation

SafetyDGX agent

arXiv:2603.18494v2 Announce Type: replace Abstract: Memory-augmented robotic policies are essential in handling memory-dependent tasks. However, existing approaches typically rely on simply extending

Minute-Scale Training for Microrobot Navigation

Model ReleasesDGX agent

arXiv:2608.00854v1 Announce Type: new Abstract: Microrobots hold significant potential for various applications, where targeted navigation is a basic requirement. Deep reinforcement learning (DRL) has

MixedComplementarityProblems.jl: A Fast, Batched, Open-Source Interior Point Solver for Mixed Complementarity Problems

Model ReleasesDGX agent

arXiv:2608.00959v1 Announce Type: cross Abstract: Mixed complementarity problems (MCPs) arise as the first-order optimality conditions of nonlinear programs and noncooperative games, and provide a nat

Motion Planning for Mobile Manipulators Navigating Doorways via Model Predictive Control

ResearchDGX agent

arXiv:2608.00206v1 Announce Type: new Abstract: Navigating doorways is a fundamental capability for mobile manipulators operating in human environments, requiring coordinated motion between the mobile

Multi-View Unified Camera Fields: Geometry-Shaped Action-Facing Representations for RGB-Only Multi-Camera VLA Policies

ApplicationsDGX agent

arXiv:2608.01826v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong generalization in robotic manipulation, yet complex contact-rich tasks often benefit from multi-ca

OC-VLA++: Monocular Geometry-Guided Cross-View Consistency for Viewpoint-Robust Robotic Manipulation

ApplicationsDGX agent

arXiv:2608.01066v1 Announce Type: new Abstract: We propose OC-VLA++, an extension of OC-VLA for viewpoint generalization under limited camera coverage. While OC-VLA grounds robot actions in the camera

OmniAI: A Surface-Adaptive Aerial Projection Interface for Human--Drone Interaction

AgentsDGX agent

arXiv:2608.00721v1 Announce Type: new Abstract: Drones in human environments often lack spatially grounded in- terfaces for situated communication. We present OmniAI, an em- bodied aerial agent that s

Perception-and-action system for humanoid robot task execution in construction

TutorialsDGX agent

arXiv:2608.01600v1 Announce Type: new Abstract: Humanoid robots, with their human-like shape and multi-tasking capabilities, are well-aligned with human-dominated workplaces, like those in civil and c

Probabilistic Reachable-Action Verification of Visuomotor Policies via Set-Based Training

SafetyDGX agent

arXiv:2608.02545v1 Announce Type: new Abstract: Reachability analysis for visuomotor policies is difficult because large visual encoders make end-to-end set propagation computationally expensive and e

← Previous
1…34567…58
Next →