Kairos: A Scalable Serving System for Physical AI
arXiv:2605.11381v1 Announce Type: new Abstract: Physical AI is experiencing rapid growth with frontier foundation models increasing its capabilities across general environments. Physical AI tasks are
Knowledge catalogue
arXiv:2605.11381v1 Announce Type: new Abstract: Physical AI is experiencing rapid growth with frontier foundation models increasing its capabilities across general environments. Physical AI tasks are
arXiv:2605.11832v1 Announce Type: new Abstract: This paper tackles spatial perception and manipulation challenges in Vision-Language-Action (VLA) models. To address depth ambiguity from monocular inpu
arXiv:2603.23679v2 Announce Type: replace Abstract: Agriculture remains a cornerstone of global health and economic sustainability, yet labor-intensive tasks such as harvesting high-value crops contin
arXiv:2605.11825v1 Announce Type: new Abstract: Affective touch in human-robot interaction is shaped not only by emotional intent, but also by robot embodiment, including touch location, physical cons
arXiv:2605.12228v1 Announce Type: new Abstract: Mobile manipulation requires coordinated control of high-dimensional, bimanual robots. Imitation learning methods have been broadly used to solve these
arXiv:2605.11665v1 Announce Type: new Abstract: Robot learning research is fragmented across policy families, benchmark suites, and real robots; each implementation is entangled with the others in a c
arXiv:2605.11762v1 Announce Type: new Abstract: Learning robust navigation policies remains a core challenge in robotics. Offline imitation learning suffers from distribution shift and compounding err
arXiv:2605.11479v1 Announce Type: new Abstract: Policy evaluation is a fundamental component of the development and deployment pipeline for robotic policies. In modern manipulation systems, this probl
arXiv:2605.12160v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies are typically evaluated as if the user had finished typing or speaking before the robot begins acting. In real dep
arXiv:2605.11534v1 Announce Type: new Abstract: When an LLM-based embodied agent fails at a household task, the culprit could be misidentified objects, forgotten sub-goals, or poor action sequencing -
arXiv:2605.11697v1 Announce Type: new Abstract: This paper presents a kinematics-aware deep reinforcement learning framework based on Rainbow Deep Q-Networks (DQN) for cooperative peg-in-hole manipula
arXiv:2605.11151v1 Announce Type: cross Abstract: Offline-to-online reinforcement learning (RL) improves sample efficiency by leveraging pre-collected datasets prior to online interaction. A key chall
arXiv:2605.12347v1 Announce Type: new Abstract: Stable, low-latency whole-body teleoperation of humanoid robots is an open research challenge, complicated by kinematic mismatches between human and rob
arXiv:2508.21260v2 Announce Type: replace Abstract: Many estimation problems in aerospace navigation and robotics involve measurements that depend on prior states. A prominent example is odometry, whi
arXiv:2605.11564v1 Announce Type: new Abstract: Despite recent efforts to collect multi-task, multi-embodiment datasets, to design recipes for training Vision-Language-Action models (VLAs), and to sho
arXiv:2605.12059v1 Announce Type: cross Abstract: Computational thinking (CT) is increasingly promoted as a core literacy, yet learners and teachers face challenges in connecting abstract program logi
arXiv:2404.05120v2 Announce Type: replace Abstract: Spherical robots typically require at least two actuators to achieve controlled 2D planar motion. Here we present Rollbot, the first spherical robot
arXiv:2605.12386v1 Announce Type: new Abstract: Robotic manipulation is typically evaluated by task success, but successful completion does not guarantee safe execution. Many safety failures are tempo
arXiv:2512.00775v2 Announce Type: replace Abstract: Linear Temporal Logic (LTL) provides a rigorous framework for specifying long-horizon robotic tasks, yet existing approaches face a trade-off: model
arXiv:2605.11618v1 Announce Type: new Abstract: Follow-the-leader (FTL) motion exploits the unique morphology of continuum robots (CRs) to navigate confined spaces by having the body retrace the path
arXiv:2605.11114v1 Announce Type: new Abstract: Vision-Language-Action (VLA) and imitation-learning policies trained via community toolchains on low-cost hardware frequently fail when deployed outside
arXiv:2605.12247v1 Announce Type: new Abstract: Contact-rich assembly is fundamental in robotics but poses significant challenges due to uncertainties in relative poses, such as misalignments and smal
arXiv:2506.14097v2 Announce Type: replace Abstract: This paper reformulates complementarity-based time-stepping for frictionless nonsmooth contact between smooth rigid bodies as a recursively generate
arXiv:2602.21625v2 Announce Type: replace Abstract: Vision-Based Tactile Sensors (VBTS) are essential for achieving dexterous robotic manipulation, yet the tactile sim-to-real gap remains a fundamenta
arXiv:2603.11383v2 Announce Type: replace Abstract: Teleoperation of low-cost robotic manipulators remains challenging due to the difficulty of retargeting human hand motion to robot joint commands. W
arXiv:2605.12162v1 Announce Type: new Abstract: Effectively handling the interplay between spatial perception and action generation remains a critical bottleneck in robotic manipulation. Existing meth
arXiv:2605.10086v1 Announce Type: new Abstract: This paper proposes a cell decomposition algorithm for binary occupancy grids that ensures mutual complete visibility from each cell to at least one adj
arXiv:2605.08947v1 Announce Type: new Abstract: This paper introduces a low-cost experimental mockup to simulate the laser cutting process of containers in nuclear decommissioning. It is composed of a
arXiv:2510.19407v2 Announce Type: replace-cross Abstract: A sensor has the ability to probe its surroundings. However, uncertainties in its exact location can significantly compromise its sensing perf
arXiv:2605.08757v1 Announce Type: new Abstract: We present a visuo-tactile data-collection system that generates temporally structured, contact-rich demonstrations for imitation learning. Conventional
arXiv:2605.09811v1 Announce Type: new Abstract: Multi-robot simultaneous localization and mapping (SLAM) is a fundamental task in multi-robot operations. Robots must have a common understanding of the
arXiv:2605.06042v2 Announce Type: replace Abstract: Flapping-wing micro aerial vehicles offer quieter and safer operation than rotary-wing drones, yet achieving precise autonomous control of bird-scal
arXiv:2605.08269v1 Announce Type: new Abstract: Wireless capsule endoscopy (WCE) enables painless visualization of the gastrointestinal tract, but its diagnostic potential is limited by incomplete muc
arXiv:2605.09659v1 Announce Type: new Abstract: Koopman operator theory provides a powerful framework for representing nonlinear dynamics through a linear operator acting on lifted observables, enabli
arXiv:2605.08831v1 Announce Type: new Abstract: In flexible assembly systems, existing task planning methods require a time-consuming configuration process by multiple experts to establish a productio
arXiv:2605.08612v1 Announce Type: new Abstract: Addressing the escalating security vulnerabilities in Vision-Language-Action (VLA) models, this study investigates backdoor attacks targeting the visual
arXiv:2605.08571v1 Announce Type: new Abstract: We introduce BEACON--Best-Effort Adaptation for Cross-Domain Co-Training--a theory-driven framework for training generative robot policies with abundant
arXiv:2605.09441v1 Announce Type: new Abstract: The pursuit of general-purpose embodied agents is hindered by fragmented evaluation protocols that isolate navigation skills and fixate on specific robo
arXiv:2605.10034v1 Announce Type: new Abstract: Recent Autonomous Driving (AD) works such as GigaFlow and PufferDrive have unlocked Reinforcement Learning (RL) at scale as a training strategy for driv
arXiv:2601.23087v3 Announce Type: replace Abstract: Learning long-horizon robotic manipulation requires jointly achieving expressive behavior modeling, real-time inference, and stable execution, which
arXiv:2605.08804v1 Announce Type: new Abstract: Reinforcement learning combined with imitation learning has significantly advanced biomimetic quadrupedal locomotion. However, scaling these frameworks
arXiv:2602.07209v2 Announce Type: replace Abstract: Localization and mapping of an environment are crucial tasks for any robot operating in unstructured environments. Time-of-flight (ToF) sensors (e.g
arXiv:2605.09216v1 Announce Type: new Abstract: Predicting the shape of tendon driven continuum robots (TDCRs) at steady state from actuation remains challenging due to continuous deformation, complex
arXiv:2503.03481v2 Announce Type: replace Abstract: This work demonstrates that the non-stop flights of three or more carriers are compatible with holding a constant pose of a cable-suspended load. It
arXiv:2605.10166v1 Announce Type: new Abstract: Robotic imitation learning typically assumes access to optimal demonstrations, yet real-world data collection often yields suboptimal, exploratory, or e
arXiv:2605.10738v1 Announce Type: cross Abstract: Decentralized collision avoidance remains challenging, particularly when agents do not communicate any information related to planned trajectories. Mo
arXiv:2507.01008v3 Announce Type: replace Abstract: Development of dexterous manipulation hardware has primarily focused on hands and grippers. However, these end-effectors are often paired with bulky
arXiv:2605.09537v1 Announce Type: new Abstract: Despite rapid progress in Vision-Language-Action (VLA) models for robotic control, instruction drift remains a persistent failure mode in long-horizon t
arXiv:2605.09801v1 Announce Type: new Abstract: Solving multi-robot motion planning (MRMP) requires generating collision-free kinodynamically feasible trajectories for multiple interacting robots. We
arXiv:2605.10063v1 Announce Type: new Abstract: Learning dynamic whole-body motions for legged robots through reinforcement learning (RL) remains challenging due to the high risk of failure, which mak
arXiv:2605.08799v1 Announce Type: new Abstract: Diffusion policies have demonstrated exceptional performance in embodied AI. However, their iterative denoising process results in high latency, and exi
arXiv:2605.10653v1 Announce Type: new Abstract: Embodied artificial intelligence is rapidly moving from research into real-world systems such as autonomous vehicles, mobile robots, and industrial mach
arXiv:2411.05516v3 Announce Type: replace Abstract: Autonomous Underwater Vehicles (AUVs) have advanced significantly in obstacle detection and path planning through sonar, cameras, and learning-based
arXiv:2511.18374v2 Announce Type: replace Abstract: We derive a computable closed-form upper bound on the Hausdorff distance between a truncated minimal robust positively invariant (mRPI) set and its
arXiv:2605.09944v1 Announce Type: new Abstract: Robust humanoid stair climbing remains challenging due to geometric discontinuities, sensitivity to step height variations, and perception uncertainty i
arXiv:2605.08434v1 Announce Type: new Abstract: Vision-language-action (VLA) models provide a promising paradigm for scalable robotic manipulation, yet their reliance on success-only behavioral clonin
arXiv:2603.16231v2 Announce Type: replace-cross Abstract: Numerical optimal control has long been split between globally structured but dimensionally intractable Hamilton--Jacobi--Bellman (HJB) method
arXiv:2602.22088v2 Announce Type: replace Abstract: Contact-rich manipulation demands human-like integration of perception and force feedback: vision should guide task progress, while high-frequency i
arXiv:2503.12333v2 Announce Type: replace Abstract: Safe, agile, and socially compliant multi-robot navigation in cluttered and constrained environments remains a critical challenge. This is especiall
arXiv:2605.10457v1 Announce Type: cross Abstract: Real-time Light Detection And Ranging (LiDAR) simulation must find, per emitted ray, the closest intersecting triangle even in dynamic scenes containi