Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
4787 storiesRaspi$^2$USBL: An open-source Raspberry Pi-Based Passive Inverted Ultra-Short Baseline Positioning System for Underwater Robotics
arXiv:2511.06998v2 Announce Type: replace Abstract: Precise underwater positioning remains a fundamental challenge for underwater robotics because global naviga…
SimCoachCorpus: A naturalistic dataset with language and trajectories for embodied teaching
arXiv:2509.14548v2 Announce Type: replace Abstract: High-quality curated datasets are essential for training and evaluating AI approaches, but are often lacking…
One-Step Model Predictive Path Integral for Manipulator Motion Planning Using Configuration Space Distance Fields
arXiv:2509.00836v3 Announce Type: replace Abstract: Motion planning for robotic manipulators is a fundamental problem in robotics. Classical optimization-based …
PROSE: Training-Free Egocentric Scene Registration with Vision-Language Models
arXiv:2606.16569v1 Announce Type: cross Abstract: Registering two captures of the same indoor space taken at different times underpins persistent spatial memory…
Decoupled Object-Centric Video Understanding for Generating Robotic Manipulation Commands
arXiv:2606.16470v1 Announce Type: cross Abstract: Translating video demonstrations into executable robot commands remains challenging because existing methods o…
A Formal Resilience Framework for Cyber-Physical Embodied Systems under Device-Level Cyberattacks
arXiv:2606.16467v1 Announce Type: cross Abstract: In cyber-physical systems (CPSs), fault tolerance is traditionally achieved by analysing sensor and actuator o…
A Pragmatic VLA Foundation Model
arXiv:2601.18692v3 Announce Type: replace Abstract: Offering great potential in robotic manipulation, a capable Vision-Language-Action (VLA) foundation model is…
Multi-Agent Embodied Autonomous Driving: From V2X Information Exchange to Shared World Models
arXiv:2606.13840v1 Announce Type: new Abstract: Autonomous driving is shifting from isolated vehicle intelligence toward multi-agent embodied systems that share…
TRACE: Trajectory-Routed Causal Memory for Delayed-Evidence Visuomotor Imitation
arXiv:2606.14551v1 Announce Type: new Abstract: Robots under autonomous operation may require decisions based on evidence that is no longer visible. We study \e…
The N2D Haptic Glove: A Multi-Finger Glove for 2D Directional Force Feedback for Contact Rich Manipulation
arXiv:2606.14083v1 Announce Type: new Abstract: Humans rely on directional fingertip forces to probe and regulate contact during manipulation, yet most wearable…
ReactSim-Bench: Benchmarking Reactive Behavior World Model Simulation in Autonomous Driving
arXiv:2606.14058v1 Announce Type: new Abstract: Reactive capability is a key property of data-driven behavior world model simulators for autonomous driving simu…
Self-Improving VLA Policies: Selected Diffusion Noise for Spurious-Robust Action Smoothing
arXiv:2606.14084v1 Announce Type: new Abstract: Diffusion-based Vision-Language-Action (VLA) policies enable strong generalization in robotic manipulation, but …
A Modular Dual-Arm Apple Harvesting Robot with Enhanced Field Performance
arXiv:2606.14089v1 Announce Type: new Abstract: Robotic apple harvesting offers a promising solution to labor shortages in commercial orchards, but low throughp…
Estimation of Ground Reaction Forces from Kinematic Data during Locomotion
arXiv:2602.03177v2 Announce Type: replace Abstract: Ground reaction forces (GRFs) provide fundamental insight into human gait mechanics and are widely used to a…
X-Loco: Towards Generalist Humanoid Locomotion Control via Synergetic Policy Distillation
arXiv:2603.03733v2 Announce Type: replace Abstract: While recent advances have demonstrated strong performance in individual humanoid skills such as upright loc…
Selective Agentic Recovery for UAV Autonomy with a Persistent Mission Runtime
arXiv:2606.14219v1 Announce Type: new Abstract: Agentic AI can support unmanned aerial vehicle (UAV) autonomy by providing high-level recovery reasoning when lo…
$\mu_0$: A Scalable 3D Interaction-Trace World Model
arXiv:2606.13769v1 Announce Type: new Abstract: World models that capture how actions induce physical change enable scalable robot learning without reliance on …
ParkourFormer: Integrating Predictive Supervision and Sequence Modeling into Parkour Locomotion
arXiv:2605.25782v3 Announce Type: replace Abstract: Humanoid parkour requires locomotion policies to coordinate whole-body dynamics across rapidly changing terr…
Efficient Domain-Adaptive Policy Learning via Kernel Representation with Application to Quadrotor Control under Non-Stationary Disturbances
arXiv:2606.13842v1 Announce Type: new Abstract: We present an algorithm for efficient domain-adaptive policy learning via kernel representations. Learning domai…
GAIT: Legged Robot Proprioceptive State Estimation with Attention over Inertial-Leg Tokens
arXiv:2606.14160v1 Announce Type: new Abstract: In this paper, we propose a method that applies Inertial-Leg (IL) tokenization to an attention-based network for…
EqCollide: Equivariant and Collision-Aware Deformable Objects Neural Simulator
arXiv:2506.05797v2 Announce Type: replace-cross Abstract: Simulating collisions of deformable objects is a fundamental yet challenging task due to the complexit…
Robustness without Wrinkles: Parallel Simulation and Robust MPC for Certified Deformable Manipulation
arXiv:2606.14188v1 Announce Type: new Abstract: We present CORD-SLS, a real-time control method for safe deformable object manipulation, with a focus on ropes a…
BIM-Loc: BIM-Integrated Discrepancy-Aware LiDAR-based Indoor Localization
arXiv:2606.14237v1 Announce Type: new Abstract: Accurate and robust localization is a fundamental requirement for service and inspection robots, particularly in…
When and How Severely: Scenario-Specific Safety Envelopes for Driving VLAs
arXiv:2606.14238v1 Announce Type: new Abstract: Safety certification of Vision-Language-Action (VLA) driving planners under ISO 21448 (SOTIF) rests on an Operat…
SyLink Hand: A Synergy-Inspired Linkage-Driven Anthropomorphic Hand for Human-Like Dexterity
arXiv:2606.14250v1 Announce Type: new Abstract: Designing anthropomorphic robotic hands that balance functional dexterity with mechanical simplicity remains a s…
Planning with the Views via Scene Self-Exploration
arXiv:2605.29563v2 Announce Type: replace-cross Abstract: Can VLMs predict how each camera move changes the view, and plan many such moves ahead? We call this c…
Spatially Conditioned Diffusion Policy: Learning Precise and Robust Manipulation with a Single RGB Camera
arXiv:2606.14535v1 Announce Type: new Abstract: Recent visual imitation learning systems have widely adopted multi-camera setups with wrist-mounted cameras as t…
ContactWorld: What Matters in Vision-Tactile World Models for Contact-Rich Manipulation
arXiv:2606.13877v1 Announce Type: new Abstract: Contact-rich manipulation requires world models to reason over complex contact dynamics from multimodal sensory …
Occupancy-Grounded Room Segmentation for Hierarchical 3D Scene Graphs
arXiv:2606.13727v1 Announce Type: new Abstract: Hierarchical 3D scene graphs (3DSGs) for indoor robots organize geometric and semantic information across spatia…
An Attention-based Model for Robust Forecasting with Missing Modality
arXiv:2606.13970v1 Announce Type: new Abstract: Learning with missing modalities is a fundamental challenge in multimodal robot learning, as real-world robotic …