Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
4787 storiesTowards safe and optimal flight: Viability Kernel MPC for Fully Actuated Multirotor
arXiv:2608.25459v1 Announce Type: new Abstract: Industrial aerial robotics demands safety guarantees for navigation in unstructured environments while optimizin…
Latent Chain-of-Thought World Modeling for End-to-End Driving
arXiv:2512.10226v3 Announce Type: replace-cross Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving explore inference-time reasoning as …
Regularized Latent Dynamics Prediction is a Strong Baseline For Behavioral Foundation Models
arXiv:2603.15857v2 Announce Type: replace-cross Abstract: Behavioral Foundation Models (BFMs) produce agents with the capability to adapt to any unknown reward …
Minimalist Visual Inertial Odometry
arXiv:2605.19990v2 Announce Type: replace Abstract: Visual-Inertial Odometry (VIO), which is critical to mobile robot navigation, uses cameras with a large numb…
UNRIO: Uncertainty-Aware Velocity Learning for Radar-Inertial Odometry
arXiv:2604.13584v2 Announce Type: replace Abstract: mmWave radars are robust to darkness and occlusions such as dust and smoke, and can directly constrain ego-v…
Scene2Demo: Self-Evolving Embodied Data Generation via Object-Action Graph
arXiv:2602.12065v2 Announce Type: replace Abstract: We present Scene2Demo, a self-evolving framework for offline embodied data generation. Given a single real-w…
MI-SLAM: Magnetic Inertial SLAM Systems
arXiv:2512.10128v4 Announce Type: replace Abstract: Spatially inhomogeneous magnetic fields offer a valuable, non-visual information source for positioning. Amo…
Choose Your Game Wisely: Measuring Game-Theoretic Structures in Real-World Vehicle Interactions
arXiv:2608.25917v1 Announce Type: cross Abstract: Game-theoretic models provide principled frameworks for modeling vehicle interactions, but their underlying te…
Saliency-Depth Conditioning for Zero-Shot Segmentation of Communication-Tower Components in Cluttered UAV Imagery
arXiv:2608.25435v1 Announce Type: cross Abstract: Fine-grained segmentation of communication-tower components in UAV imagery is essential for automated inspecti…
PhaseShift: Topology-Aware Data Harmonization and Model Consolidation Across Signalized Intersections
arXiv:2608.25275v1 Announce Type: cross Abstract: Learned traffic-behavior models are commonly trained separately for each intersection, creating model portfoli…
Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization
arXiv:2608.26103v1 Announce Type: new Abstract: Zero-shot cross-task generalization, where a policy must execute manipulation tasks never seen during training, …
Fast Generative Grasping via Lie Group-Constrained MeanFlow
arXiv:2608.26076v1 Announce Type: new Abstract: Grasp synthesis is a core task in robotic manipulation, for which the solution typically forms a multimodal dist…
Phantom Navigator: Stealthy and Precise Unmanned Aerial Vehicle Redirection with Real-Time Tracking and GPS Spoofing
arXiv:2608.26011v1 Announce Type: new Abstract: Redirecting unmanned aerial vehicles (UAVs) from their intended mission trajectories has been an active area of …
A Statistical Audit of Physical AI Benchmark Redundancy
arXiv:2608.25940v1 Announce Type: new Abstract: Physical AI models are evaluated on suites of benchmarks that differ across model reports, leaving the model-by-…
TacForcing: Streaming Action Generation with Execution-Time Tactile Feedback
arXiv:2608.25798v1 Announce Type: new Abstract: Contact-rich manipulation requires adapting to contact states that can evolve substantially within an action hor…
LM-X: Explainable Action Modeling with Progress, Event, and Uncertainty Prediction for Generalist Robot Manipulation
arXiv:2608.25757v1 Announce Type: new Abstract: Generalist vision--language--action (VLA) policies learn long-horizon behavior mainly through short-horizon acti…
PRISM: Projection-Integrated Sampling-Based MPC with Bayesian Cost Tuning for Bimanual Manipulation
arXiv:2608.25666v1 Announce Type: new Abstract: Bimanual manipulation in cluttered, contact-rich environments remains challenging because it requires coordinate…
GaussianDream++: Efficient 3D Gaussian World Modeling for Robotic Manipulation
arXiv:2608.25659v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies have advanced language-conditioned robotic manipulation, yet action-imitat…
RA-VLA: Retrieval-Augmented VLA for Test-Time Adaptation
arXiv:2608.25585v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models provide a versatile foundation for general robotic manipulation, yet they ex…
Transient multimode heat transfer of an industrial automated tape laying process under rapidly changing conditions
arXiv:2608.25470v1 Announce Type: new Abstract: This work presents a transient heat-transfer model of an industrial automated tape laying (ATL) process designed…
A Taxonomy of Construction Task Activities for Robot Workers
arXiv:2608.25395v1 Announce Type: new Abstract: Recent vision-language-action models offer a path toward robots with broader repertoires than conventional task-…
RAEM: Robust Autonomous Exploration for Multi-Floor Environments with a Quadruped Robot
arXiv:2608.25366v1 Announce Type: new Abstract: In this paper, we propose RAEM, a robust autonomous exploration framework for quadruped robots operating in mult…
Development of a Voice-Controlled Tendon-Driven Bionic Hand
arXiv:2608.25222v1 Announce Type: new Abstract: The impairment of the hands can seriously affect the abilities of every individual to perform the every-day acti…
ROS2 Connect: A new ROS2 over WAN Solution
arXiv:2608.25102v1 Announce Type: new Abstract: The Robot Operating System 2 (ROS2) has become a widely adopted framework for the development of distributed rob…
Gatik brings in $200M to continue expanding autonomous trucking operations
Gatik said the funding will help it expand its model built around high-frequency regional routes connecting distribution centers and stores. The post Gatik brin…
CARE: Camera-Residual Reserves for First Sightings in Adaptive LiDAR Sensing
arXiv:2608.24282v1 Announce Type: cross Abstract: Adaptive LiDAR scanning concentrates a limited sensing budget on regions of interest predicted from past objec…
A study on the effects of mixed explicit and implicit communications in human-artificial-agent interactions
arXiv:2409.18745v5 Announce Type: replace Abstract: Communication between humans and artificial agents is essential for their interaction. This is often inspire…
VGGT-DP: Generalizable Robot Control via Vision Foundation Models
arXiv:2509.18778v2 Announce Type: replace Abstract: Visual imitation learning frameworks allow robots to learn manipulation skills from expert demonstrations. W…
A Robust Task-Level Control Architecture for Learned Dynamical Systems
arXiv:2511.09790v2 Announce Type: replace Abstract: Dynamical system (DS)-based learning from demonstration (LfD) is a powerful tool for generating motion plans…
E2HiL: Entropy-Guided Sample Selection for Efficient Real-World Human-in-the-Loop Reinforcement Learning
arXiv:2601.19969v2 Announce Type: replace Abstract: Human-in-the-loop guidance has emerged as an effective approach for accelerating online reinforcement learni…