Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
2771 storiesWatchAct: A Benchmark for Behavior-Grounded Robot Manipulation
arXiv:2606.26443v1 Announce Type: new Abstract: A robot working alongside people must reason about what they have done, in what order, and with what intent. Vid…
IDEA: Insensitive to Dynamics Mismatch via Effect Alignment for Sim-to-Real Transfer in Multi-Agent Control
arXiv:2606.26575v1 Announce Type: new Abstract: Complex multi-agent control tasks remain challenging for traditional rule-based and model-based approaches, moti…
CogAD: Cognitive-Hierarchy Guided End-to-End Autonomous Driving
arXiv:2505.21581v4 Announce Type: replace Abstract: While end-to-end autonomous driving has advanced significantly, prevailing methods remain fundamentally misa…
ForesightSafety-VLA: A Unified Diagnostic Safety Benchmark for Vision-Language-Action Models
arXiv:2606.27079v1 Announce Type: new Abstract: In embodied intelligence, safety is a prerequisite for reliable robot deployment in the physical world. Current …
Automating Potential-based Reward Shaping with Vision Language Model Guidance
arXiv:2606.27180v1 Announce Type: cross Abstract: Sparse rewards are inherently challenging for reinforcement learning agents as they lack intermediate feedback…
Hardware Design for Table Tennis Robot Capable of Beating Professional Players
arXiv:2606.26643v1 Announce Type: new Abstract: This paper focuses on the hardware specifications required for a table tennis robot to beat professional players…
E-TTS: A New Embodied Test-Time Scaling Framework for Robotic Manipulation
arXiv:2606.27268v1 Announce Type: new Abstract: Recently, a few works have made early attempts to study test-time scaling for embodied tasks. However, two major…
Morphology-Specific Closed-Loop Control of Logarithmic-Spiral Continuum Arms via Online Jacobian Error Compensation
arXiv:2606.26188v1 Announce Type: new Abstract: Logarithmic spirals are ubiquitous in biological appendages and provide an attractive morphology for continuum m…
Bridging Performance and Generalization in Reinforcement Learning for Agile Flight
arXiv:2606.27348v1 Announce Type: new Abstract: Autonomous drone racing is a fundamentally challenging regime for autonomous aerial robots, requiring time-optim…
OmniRobotHome: A Multi-Camera Home Platform for Real-Time Human-Robot Interaction
arXiv:2604.28197v2 Announce Type: replace Abstract: Robots in homes must continuously sense the people around them, yet most prior work relies on limited or off…
Real-Time Safety Evaluation of Human Arm Operations Using a Wrist-Mounted IMU with PSM System
arXiv:2502.09241v2 Announce Type: replace Abstract: This paper presents a novel approach to real-time safety monitoring in human-robot collaborative manufacturi…
LAMP: Lane-Aligned Motion Primitives for Feasible Trajectory Prediction
arXiv:2606.26661v1 Announce Type: new Abstract: Motion forecasting is essential for autonomous driving systems to enable safe decision-making and planning in co…
FlameVQA: A Physically-Grounded UAV Wildfire VQA Benchmark with Radiometric Thermal Supervision
arXiv:2606.27128v1 Announce Type: cross Abstract: Wildfire monitoring from UAVs requires reliable reasoning over complex aerial scenes, where smoke, scale varia…
PRISM: Efficient and Locally Optimal Probabilistic Planning with Reachability Guarantees
arXiv:2606.26413v1 Announce Type: cross Abstract: Belief-space planning under motion uncertainty and state and control constraints remains a fundamental challen…
Reinforcement Learning Enables Autonomous Microrobot Navigation and Intervention in Simulated Blood Capillaries
arXiv:2606.26154v1 Announce Type: new Abstract: Autonomous microrobots navigating biological vasculature could enable targeted drug delivery and thrombolysis, y…
History-Conditioned Spatio-Temporal Visual Token Pruning for Efficient Vision-Language Navigation
arXiv:2603.06480v2 Announce Type: replace Abstract: Vision-Language Navigation (VLN) enables robots to follow natural-language instructions in visually grounded…
CoStream: Composing Simple Behaviors for Generalizable Complex Manipulation
arXiv:2606.26423v1 Announce Type: new Abstract: Long-horizon, contact-rich complex manipulation tasks, such as seating a GPU into a PCIe slot, demand both milli…
Unsupervised Memory-Enhanced Video Transformers: Obstacle Detection for Autonomous Agricultural Rover
arXiv:2606.26151v1 Announce Type: new Abstract: While autonomous rovers have become indispensable to precision farming, achieving consistent operational safety …
ForceBand: Learning Forceful Manipulation with sEMG
arXiv:2606.26093v1 Announce Type: new Abstract: Human demonstrations are a scalable data source for learning robot manipulation policies. However, common source…
Memory-Efficient Policy Libraries with Low-Rank Adaptation in Reinforcement Learning
arXiv:2606.25700v1 Announce Type: cross Abstract: When fine-tuning Large Language Models (LLMs), there has been success in minimizing both memory usage and comp…
TIDAL: Temporally Interleaved Diffusion and Action Loop for High-Frequency VLA Control
arXiv:2601.14945v2 Announce Type: replace Abstract: Large-scale Vision-Language-Action (VLA) models offer semantic generalization but suffer from high inference…
Calousel: Extrinsic Calibration of Non-overlapping Multi-camera Systems from Pure Rotation
arXiv:2606.25646v1 Announce Type: new Abstract: Extrinsic calibration of multi-camera systems with non-overlapping FOVs has been a challenging problem in the ro…
RARM: Confidence-Gated Progress Reward Modeling for RL in Manipulation
arXiv:2606.22027v2 Announce Type: replace Abstract: Reinforcement learning for robot manipulation is often bottlenecked by reward design, especially in long-hor…
Learning Robot Visual Navigation in Crowds via Intention-Aware Scene Representations
arXiv:2606.26047v1 Announce Type: new Abstract: Robot crowd navigation requires the ability to infer human intentions while accounting for the structural constr…
Large-Scale Tunnel Air--Ground Collaboration With FLISP: Fast LiDAR-IMU Synchronized Path Planne
arXiv:2606.25393v1 Announce Type: new Abstract: Hydropower tunnel inspection is critical for infrastructure integrity yet remains inefficient and hazardous usin…
RoboRouter: Training-Free Policy Routing for Robotic Manipulation
arXiv:2603.07892v4 Announce Type: replace Abstract: Research on robotic manipulation has developed a diverse set of policy paradigms, including vision-language-…
FAR-LIO: Enabling High-Speed Autonomy through Fast, Accurate, and Robust LiDAR-Inertial Odometry
arXiv:2606.26010v1 Announce Type: new Abstract: Robust and accurate odometry estimation is essential in modern robotics. In environments characterized by highly…
Reasonable Motion: A General ASP Foundation for Environment Constrained Movement Trajectory Computation
arXiv:2606.25626v1 Announce Type: cross Abstract: We present a general answer set programming based hybrid quantitative-qualitative method for computing constra…
Spatio-Temporal Retrieval-based Priors for Adaptive Computational Teaching in Driving
arXiv:2606.25224v1 Announce Type: new Abstract: Learning-based automated coaching systems for complex motor tasks such as high-performance driving remain limite…
SwarmFly: A simulation platform for UAV swarm experiment design and validation
arXiv:2606.25146v1 Announce Type: new Abstract: The initial development phase of UAV swarms largely depends on simulation for experimental design and validation…