Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
2771 storiesPlay2Perfect: What Matters in Dexterous Play Pretraining for Precise Assembly?
arXiv:2606.26428v1 Announce Type: new Abstract: Multi-fingered robots promise the speed and dexterity of human hands, yet challenging problems such as precise a…
Monte Carlo Tree Search with Tensor Factorization for Optimization Problems in Robotics
arXiv:2507.04949v3 Announce Type: replace Abstract: Many robotic tasks, such as inverse kinematics, motion planning, and contact-rich manipulation, can be formu…
A System for Fast, Resilient, and Adaptable Loco-Manipulation Behaviors on Humanoid Robots
arXiv:2606.26425v1 Announce Type: new Abstract: Humanoid robots could take on physically demanding, hazardous, and repetitive work in spaces built for humans. H…
FC-Vision: Real-Time Visibility-Aware Replanning for Occlusion-Free Aerial Target Structure Scanning in Unknown Environments
arXiv:2602.13720v2 Announce Type: replace Abstract: Autonomous aerial scanning of target structures is crucial for practical applications, requiring online adap…
CycleRL: Sim-to-Real Deep Reinforcement Learning for Robust Autonomous Bicycle Control
arXiv:2603.15013v3 Announce Type: replace Abstract: Autonomous bicycles offer a promising agile solution for urban mobility and last-mile logistics. However, co…
Tactile-WAM: Touch-Aware World Action Model with Tactile Asymmetric Attention
arXiv:2606.26663v1 Announce Type: new Abstract: World Action Models (WAMs) generate actions together with predicted futures, offering a powerful interface for r…
Layered Outer-Loop Control for Disturbance-Robust Multi-Waypoint UAV Arrival
arXiv:2606.26315v1 Announce Type: new Abstract: Disturbance-robust UAV position control is easy to demonstrate in benign simulations but much harder to make fas…
PressMimic: Pressure-Guided Motion Capture and Control for Humanoid Robot Imitation
arXiv:2606.26741v1 Announce Type: new Abstract: Humanoid motion imitation requires not only accurate perception of human kinematics but also faithful reproducti…
VibeAct: Vibration to Actions for Contact-Rich Reactive Robot Dexterity
arXiv:2606.27344v1 Announce Type: new Abstract: Dexterous manipulation depends on contact events that are fast, local, and often visually occluded. Piezoelectri…
KRVF: A Source-Aware Semantic Voxel World Representation for Edge Mobile Manipulation
arXiv:2606.26321v1 Announce Type: new Abstract: Mobile manipulators need world models that are current, queryable, semantically meaningful, and usable under edg…
BAT-Nav: Budget-Aware Arbitration and Termination for Long-Horizon Semantic Navigation
arXiv:2605.16932v2 Announce Type: replace Abstract: Long-horizon semantic navigation asks a robot to localize multiple open-vocabulary targets under a finite ac…
UAV-MapFusion: RTK-Aligned Uncertainty-Aware Coarse-to-Fine Multi-Session UAV Mapping
arXiv:2606.26928v1 Announce Type: new Abstract: Large-scale point cloud maps are essential for robotics and spatial intelligence tasks. UAVs provide an efficien…
NavIsaacLab: Generating Realistic Crowd via Parallel Robot Learning for Benchmarking Human-aware Navigation
arXiv:2606.26265v1 Announce Type: new Abstract: Robot autonomous navigation that accounts for surrounding human activities is crucial for ensuring both safety a…
How Should a Simulation-to-Reality Transfer Budget Be Spent?
arXiv:2606.22062v2 Announce Type: replace Abstract: Simulation-to-reality transfer, often called sim-to-real transfer, is a central challenge in robot learning.…
OmniContact: Chaining Meta-Skills via Contact Flow for Generalizable Humanoid Loco-Manipulation
arXiv:2606.26201v1 Announce Type: new Abstract: Learning long-horizon humanoid loco-manipulation poses a dual challenge: it requires not only the robust executi…
Racing a Wheeled Quadruped: Active Load Transfer Mitigation via Model Predictive Control
arXiv:2606.26313v1 Announce Type: new Abstract: This paper presents a hierarchical control framework using model predictive control (MPC) and reinforcement lear…
TaskNPoint: How to Teach Your Humanoid to Hit a Backhand in Minutes
arXiv:2606.26215v1 Announce Type: new Abstract: How do we learn to hit a tennis backhand? Not from a thousand hours of tennis tournaments on TV - we work with a…
Scalable Behavior Cloning with Open Data, Training, and Evaluation
arXiv:2606.27375v1 Announce Type: new Abstract: We introduce ABC, a fully open-source stack for manipulation with behavior cloning. At its core is ABC-130K: the…
Learning Motion Feasibility from Point Clouds in Cluttered Environments
arXiv:2606.26700v1 Announce Type: new Abstract: Motion feasibility prediction plays a central role in robotics, particularly in task and motion planning and man…
Inference-Time Robot Behavior Steering through Physically-Aware Reconfiguration of Task-Structure
arXiv:2606.26588v1 Announce Type: new Abstract: A central challenge in deploying learned robot policies is inference-time behavior steering: redirecting a polic…
BOWConnect: Parallel Bayesian Optimization over Windows with Learned Local Cost Maps for Sample-Efficient Kinodynamic Motion Planning
arXiv:2606.27292v1 Announce Type: new Abstract: This paper presents BOWConnect, a bidirectional parallel kinodynamic motion planner that addresses three fundame…
Improving Vision-Language-Action Model Fine-Tuning with Structured Stage and Keyframe Supervision
arXiv:2606.26801v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown strong potential for generalizable robotic manipulation. During f…
LiMoDE: Rethinking Lifelong Robot Manipulation from a Mixture-of-Dynamic-Experts Perspective
arXiv:2606.26183v1 Announce Type: new Abstract: Building a generalist robot that can leverage prior knowledge for continuous task adaptation remains a significa…
RoboTales: ROBOTic Anthropomorphic LEarning Systems
arXiv:2606.26213v1 Announce Type: new Abstract: RoboTales is a low-cost robotic storytelling system that animates narratives using expressive sock puppetry. Imp…
LA4VLA: Learning to Act without Seeing via Language-Action Pretraining
arXiv:2606.27295v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are commonly pretrained on robot demonstrations by jointly mapping visual ob…
GO: The Great Outdoors Multimodal Dataset
arXiv:2501.19274v3 Announce Type: replace Abstract: The Great Outdoors (GO) dataset is a multi-modal annotated data resource aimed at advancing ground robotics …
Reinforcement Fine-Tuning of Flow-Matching Policies for Vision-Language-Action Models
arXiv:2510.09976v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models such as OpenVLA, Octo, and $\pi_0$ have shown strong generalizatio…
Visual-Language-Guided Task Planning for Horticultural Robots
arXiv:2601.11906v2 Announce Type: replace Abstract: Crop monitoring is essential for precision agriculture, but current systems lack high-level reasoning. We in…
Ordinal Neural Collapse as a Representation Prior for Visual Navigation
arXiv:2606.26839v1 Announce Type: new Abstract: Learning robust navigation policies directly from visual observations remains a fundamental challenge in vision-…
HumanoidUMI: Bridging Robot-Free Demonstrations and Humanoid Whole-Body Manipulation
arXiv:2606.27239v1 Announce Type: new Abstract: High-quality demonstration data are essential for humanoid robot skill learning, especially for whole-body behav…