Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
4787 storiesASVSim (AirSim for Surface Vehicles): A High-Fidelity Simulation Framework for Autonomous Surface Vehicle Research
arXiv:2506.22174v3 Announce Type: replace Abstract: The transport industry has recently shown significant interest in unmanned surface vehicles (USVs), specific…
Homotopy-Aware Corridor Generation without Predefined Reference Paths
arXiv:2607.29513v1 Announce Type: new Abstract: Generating safe corridors is essential for collision-free robotic motion planning, yet most existing methods rel…
Balancing of Humanoid with Object Mass: Trade-off Analyses and Lifting Control
arXiv:2607.29625v1 Announce Type: new Abstract: The demand for humanoid loco-manipulation tasks with an object has recently increased, and most existing control…
TacPrint: A Wearable Fingertip Tactile Sensor for Human-to-Robot Contact Reproduction
arXiv:2607.29231v1 Announce Type: new Abstract: Human-centric data collection is emerging as a significant paradigm for robot skill acquisition, but seamlessly …
TRACT: Temporally Routed Action Chunks with Chronological Phase Authority for Contact-Rich Manipulation
arXiv:2607.29285v1 Announce Type: new Abstract: Action chunking shortens the effective decision horizon of robot imitation learning by predicting multiple futur…
ActFovea: Runtime Safeguarding for VLA Policies via Spatiotemporal Visual-Action Consistency
arXiv:2607.29169v1 Announce Type: new Abstract: Vision-language-action (VLA) policies achieve strong performance in robotic manipulation but remain vulnerable t…
Auto-JEPA: A Latent World Model of Continuous Intent for End-to-End Autonomous Driving
arXiv:2607.29031v1 Announce Type: new Abstract: Existing autonomous-driving world models typically perform dense prediction of future videos, occupancy states, …
ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts
arXiv:2607.28993v1 Announce Type: new Abstract: World Action Models (WAMs) have emerged as a promising paradigm by jointly modeling robot actions and future vis…
Leveraging Image Generators to Address Data Scarcity: The Gen4Regen Dataset for Forest Regeneration Mapping
arXiv:2605.05627v2 Announce Type: replace-cross Abstract: Sustainable forest management relies on precise species composition mapping, yet traditional ground su…
FBFM: A Training-Free Asynchronous Feedback Mechanism for Flow-Matching in World-Action Models Execution
arXiv:2607.29235v1 Announce Type: new Abstract: Although world-action models (WAMs) enhance long-horizon robot control by predicting visual evolution before act…
CLIFT: Turning Gemini Robotics On-Device into Humanoid Specialists via Non-Invasive Closed-Loop Iterative Fine-Tuning
arXiv:2607.29172v1 Announce Type: new Abstract: While robot foundation models are growing increasingly capable, the strongest models are typically trained on pr…
DuetHOI: Language-Guided Bimanual Hand--Object Motion Generation with Articulation Planning and Contact Refinement
arXiv:2603.08390v3 Announce Type: replace Abstract: Bimanual articulated-object interaction generation requires a model to capture the evolution of object artic…
HAM-VLN: Harnessing Hierarchical Agentic Memory for Zero-Shot Vision-and-Language Navigation
arXiv:2607.29600v1 Announce Type: new Abstract: Vision-and-language navigation (VLN) enables robots to follow instructions in previously unseen environments. Re…
The Open Motion Planning Library 2.0
arXiv:2605.29301v2 Announce Type: replace Abstract: The Open Motion Planning Library (OMPL), first released in 2008, has become a cornerstone of the motion plan…
Event-Based Upper-Body Humanoid Teleoperation Under Challenging Illumination
arXiv:2607.29227v1 Announce Type: new Abstract: We present a real-time upper-body human-to-humanoid motion imitation framework driven by neuromorphic event-base…
AREA3D: Active Reconstruction Agent with Unified Feed-Forward 3D Perception and Vision-Language Guidance
arXiv:2512.05131v2 Announce Type: replace-cross Abstract: Active 3D reconstruction enables an agent to autonomously select viewpoints to efficiently obtain accu…
Temporal Policy: History-Initialized Action Generation for Robotic Learning from Demonstration
arXiv:2607.29482v1 Announce Type: new Abstract: By relying on independent couplings from uninformative Gaussian priors, standard diffusion and flow matching mod…
Mirror Learning
arXiv:2607.28737v1 Announce Type: cross Abstract: We investigate imitation learning through the lens of third-person observation and propose a framework for mir…
Shepherding UAV Swarm with Action Prediction Based on Movement Constraints
arXiv:2604.17189v3 Announce Type: replace Abstract: In this study, we propose a new sheepdog-inspired control method for a swarm of small unmanned aerial vehicl…
VSTaI: Design and Characterization of Variable-Stiffness Tactile Interfaces Based on 3D-Printed Structured Fabrics
arXiv:2607.29102v1 Announce Type: new Abstract: Realistic palpation training requires reliable rendering of soft tissue stiffness changes in real time, which is…
Bootstrapping Self-Supervised Learning of Binary Classification Using Error Bounds: A Case Study on a Robotic Insertion Task
arXiv:2607.29640v1 Announce Type: new Abstract: Flexible manufacturing requires rapid deployment of solutions and minimal setup time to remain competitive. An e…
RAPiD: Reward-Guided Consistency Distillation of Diffusion Planners for Real-Time Autonomous Driving
arXiv:2602.07339v2 Announce Type: replace-cross Abstract: Diffusion-based trajectory planners can model multi-modal driving behavior, but their iterative denois…
Physics-Embedded Neural ODEs for Learning Antagonistic Pneumatic Artificial Muscle Dynamics
arXiv:2602.23670v2 Announce Type: replace Abstract: Pneumatic artificial muscles (PAMs) enable compliant actuation for soft wearable, assistive, and interactive…
Choose What to Manipulate: Revealing Data Scaling Laws in Bounding-Box Guided Policies for Semantic Manipulation
arXiv:2602.11885v2 Announce Type: replace Abstract: Diffusion-based policies generalize poorly in semantic manipulation, a key obstacle to real-world deployment…
Towards bridging the gap: Systematic sim-to-real transfer for diverse legged robots
arXiv:2509.06342v2 Announce Type: replace Abstract: Legged robots must achieve both robust locomotion and energy efficiency to be practical in real-world enviro…
DART: Dual-Axis Airborne Reachability-Gated Torque-Reaction for Off-Road Vehicle Jumps
arXiv:2607.29011v1 Announce Type: new Abstract: Traversing crests, ledges, and ditches at high speed often launches vehicles into the air, and a mishandled land…
FCC robot ruling shines a spotlight on U.S. policy; how next-gen AI can help warehousing
Osaro’s Derek Pridmore says warehouse robots are advancing, but real-world safety and reliability still depend on specialized, layered AI systems. The post FCC …
QQWorld: Quantile-Quantile Matching for World Model Regularization
arXiv:2607.28415v1 Announce Type: cross Abstract: Latent world models enable efficient planning by predicting future states in a compact representation space, b…
RMBench: Memory-Dependent Robotic Manipulation Benchmark with Insights into Policy Design
arXiv:2603.01229v3 Announce Type: replace Abstract: Robotic manipulation policies have made rapid progress in recent years, yet most existing approaches give li…
SeedPolicy: Horizon Scaling via Self-Evolving Diffusion Policy for Robot Manipulation
arXiv:2603.05117v4 Announce Type: replace Abstract: Imitation Learning (IL) enables robots to acquire manipulation skills from expert demonstrations. Diffusion …