Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
4787 storiesEgoInfinity: A Web-Scale 4D Hand-Object Interaction Data Engine for Any-View Robot Retargeting and Video-to-Action Robot Learning
arXiv:2606.17385v1 Announce Type: new Abstract: Internet videos constitute the largest reservoir of embodied human manipulation knowledge, yet converting arbitr…
Damage Adaptation in Seconds for Architected Materials
arXiv:2606.17394v1 Announce Type: new Abstract: Adaptation to damages and in-situ physical repairs is essential for long-term robot autonomy, yet challenging ou…
Where Should Action Generation Begin? A Learnable Source Prior for Generative Robot Policies
arXiv:2606.17408v1 Announce Type: new Abstract: Generative robot policies typically begin action generation from an observation-independent standard Gaussian di…
AnnotateAnything: Automatic Annotation of 3D Assets for Robot Manipulation
arXiv:2606.17446v1 Announce Type: new Abstract: Simulation enables scalable robot data collection, but raw 3D assets provide only geometry, lacking the semantic…
Continual Online Personalization of Exoskeleton Control via Manifold-Aware Experience Replay
arXiv:2606.17455v1 Announce Type: new Abstract: Personalizing exoskeleton control remains a critical challenge for clinical users with gait disabilities. Online…
RICH-SLAM: Radar SLAM with Incremental and Continuous Hilbert Mapping
arXiv:2606.17534v1 Announce Type: new Abstract: Simultaneous localization and mapping using radar sensors has gained increasing attention due to radar's inheren…
ED3R: Energy-Aware Distributed Disaster Detection Enabled by Cooperative Robotic Agents
arXiv:2606.17739v1 Announce Type: new Abstract: Robotics are expected to support environmental monitoring and natural disaster management, where decisions must …
HumanoidArena: Benchmarking Egocentric Hierarchical Whole-body Learning
arXiv:2606.17833v1 Announce Type: new Abstract: Humanoid robots promise whole-body interaction in human-centered environments, but scalable policy learning rema…
SPARK: Low Latency Single-Camera 3D Pose Estimation for Autonomous Racing using Keypoints
arXiv:2606.17936v1 Announce Type: new Abstract: In autonomous racing, fast detection of other participants' movements is required to plan safe, collision-free t…
EAGG: Embodiment-Aligned Grasp Generation via Geometry-Aware Graph Conditioning
arXiv:2606.18092v1 Announce Type: new Abstract: Cross-end-effector grasp generation seeks a unified model that generalizes across objects and across embodiments…
WireCraft: A Simulation Benchmark for Industrial DLO Manipulation
arXiv:2606.18097v1 Announce Type: new Abstract: Deformable Linear Objects (DLOs), such as wires and cables, are central to industrial assembly. Unlike rigid obj…
Extracting Semantics: LLM-Guided Automatic Population of Robot Ontology from URDF
arXiv:2606.17073v1 Announce Type: new Abstract: While commonsense knowledge may suffice for virtual agents, embodied robots interacting with humans require grou…
Intermittent Strategic Cooperation of Two Selfish Agents on Graphs
arXiv:2606.17216v1 Announce Type: cross Abstract: We study strategic space- and time-constrained cooperation between two self-interested agents through the Inte…
Beyond Benchmarks: Continuous Edge Inference for Fine-Grained Roadside Perception
arXiv:2606.17241v1 Announce Type: cross Abstract: Continuous AI inference on resource-constrained edge hardware introduces deployment effects that are largely i…
DriveJudge: Rethinking Autonomous Driving Evaluation with Vision-Language Models
arXiv:2606.17362v1 Announce Type: cross Abstract: Autonomous driving has shifted towards end-to-end policy learning, where reliable, interpretable policy evalua…
TerraTransfer: Learning End-to-End Driving Policies Without Expert Demonstrations
arXiv:2606.17386v1 Announce Type: cross Abstract: End-to-end autonomous driving has achieved state-of-the-art performance on benchmarks and real-world deploymen…
WeaveLA: Event Driven Cross-Subtask Latent Memory Weaving for Repetitive Robot Manipulation
arXiv:2606.17463v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies have achieved remarkable single-step manipulation, yet they remain britt…
GeneralVLA-2: Geometry-Aware Reconstruction and Governed Memory for Robot Planning
arXiv:2606.17480v1 Announce Type: cross Abstract: Generalist vision-language-action systems need object-centric 3D evidence and reusable manipulation experience…
Learn to Quantify Social Interaction with Constraints for Pedestrian Walking
arXiv:2606.17897v1 Announce Type: cross Abstract: Long-term human path forecasting in crowds is critical for autonomous moving platforms (like autonomous drivin…
Memory as a Wasting Asset: Pricing Flash Endurance for Embodied Agents, and the Limits of Doing So
arXiv:2606.18144v1 Announce Type: cross Abstract: A robot's flash endurance is a non-renewable stock: every persisted write spends one of a few thousand program…
SSIL: Self-Supervised Imitation Learning for End-to-End Driving
arXiv:2308.14329v4 Announce Type: replace Abstract: In autonomous driving, the end-to-end (E2E) driving approach that predicts vehicle control signals directly …
MOCHI: Motion Enhancement of Collaborative Human-object Interactions
arXiv:2606.18243v1 Announce Type: cross Abstract: Collaborative human-object interaction shows dynamic and complex movements that require mutual anticipation an…
OpenTie: Open-vocabulary Sequential Rebar Tying System
arXiv:2509.00064v2 Announce Type: replace Abstract: Robotic practices on the construction site emerge as an attention-attracting manner owing to their capabilit…
OmniRetarget: Interaction-Preserving Data Generation for Humanoid Whole-Body Loco-Manipulation and Scene Interaction
arXiv:2509.26633v3 Announce Type: replace Abstract: A dominant paradigm for teaching humanoid robots complex skills is to retarget human motions as kinematic re…
Can Vision Foundation Models Navigate? Zero-Shot Real-World Evaluation and Lessons Learned
arXiv:2603.25937v2 Announce Type: replace Abstract: Visual Navigation Models (VNMs) promise generalizable, robot navigation by learning from large-scale visual …
A 3D Isovist World Model -- Revealing a City's Unseen Geometry and Its Emergent Cross-City Signature
arXiv:2606.03609v3 Announce Type: replace Abstract: Embodied agents that navigate cities rely on world models that predict how their surroundings will change as…
Simulating Infant First-Person Sensorimotor Experience via Motion Retargeting from Babies to Humanoids
arXiv:2604.27583v2 Announce Type: replace-cross Abstract: Motion retargeting from humans to human-like artificial agents is becoming increasingly important as h…
SimTO: A two-stage, simulation-driven topology optimization framework for bespoke soft robotic grippers
arXiv:2601.19098v2 Announce Type: replace Abstract: Soft robotic grippers are essential for grasping delicate, geometrically complex objects in manufacturing, h…
AlignDrive: Aligned Lateral-Longitudinal Planning for End-to-End Autonomous Driving
arXiv:2601.01762v3 Announce Type: replace Abstract: Practical autonomous driving requires models that generalize by reasoning through spatial-temporal possibili…
Contactless Respiratory Monitoring on Heterogeneous Mobile Robots: A Multimodal Edge-Computing Framework
arXiv:2606.17376v1 Announce Type: new Abstract: Respiratory-rate (RR) monitoring is a critical component of remote triage and victim assessment in emergency res…