Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
4787 storiesDPNet: Efficient Dead-End Prediction and Avoidance for Vision-Based UAV Navigation
arXiv:2608.16640v1 Announce Type: new Abstract: Vision-based Unmanned Aerial Vehicles (UAVs) often suffer from navigation failures in dead ends due to limited s…
Zetta $\zeta$: An Efficient Closed-Loop Embodied Harness for Self-Evolving Physical Intelligence
arXiv:2608.16590v1 Announce Type: new Abstract: Embodied agents are increasingly used to close the gap left by end-to-end policy models. Yet the agentic path ha…
Co-design of Neural and Muscle Network based on Embodied Perceptron Representation
arXiv:2608.16555v1 Announce Type: new Abstract: Recent advances in AI technologies have enabled the advanced design of complex control policies. In contrast, fo…
NebulaVLA: A Dual-Frequency Vision-Language-Action Model With Guide Action for Robotic Manipulation
arXiv:2608.16503v1 Announce Type: new Abstract: Real-world deployment of Vision-Language-Action (VLA) models is often bottlenecked by efficiency-performance tra…
Cyclops: LiDAR as a Camera That Dreams in Color
arXiv:2608.16264v1 Announce Type: new Abstract: Conventionally, robotic perception relies heavily on cameras due to the rich semantic texture they provide. Howe…
Unified Condition-Action Modeling for Accurate One-Step Action Generation
arXiv:2608.16153v1 Announce Type: new Abstract: Robot manipulation requires policies that are both accurate and efficient, as robot control must respond to chan…
ScenarioCharacterization: A Modular Toolkit for Characterizing Safety across Trajectory Datasets
arXiv:2608.16041v1 Announce Type: new Abstract: We introduce ScenarioCharacterization, an open-source framework for automated, dataset-agnostic profiling of dri…
Benchmarking Identity-Sensitive LLM Outputs for Surveillance and Security Robots
arXiv:2608.16030v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used to generate textual robot design specifications, interaction …
Revisiting Open-Loop Execution in Robotics: Toward Reactive, Higher-Performing Policies
arXiv:2608.15938v1 Announce Type: new Abstract: Action chunking --- the practice of predicting a sequence of actions and executing a prefix open-loop --- has em…
GigaBrain-0.7: Scaling Embodied Foundation Models to Emergent Capabilities with a Three-System Architecture
arXiv:2608.15875v1 Announce Type: new Abstract: Vision-language-action (VLA) models have become a dominant paradigm for generalist embodied agents, demonstratin…
ViTaR: Visuo-Tactile Residual Adaptation for Foundation VLA Manipulation
arXiv:2608.15816v1 Announce Type: new Abstract: As Vision-Language-Action (VLA) models scale toward real-world deployment, contact-rich manipulation exposes a c…
Tac4Loco: Learning Spatiotemporal Plantar Pressure Representations for Humanoid Locomotion
arXiv:2608.15766v1 Announce Type: new Abstract: Humanoid robots are expected to traverse complex terrains, where the plantar support may vary dramatically due t…
Robo-Dopamine 2.0: History-Conditioned and OOD-Aware Process Reward Modeling for Robotic Manipulation
arXiv:2608.15680v1 Announce Type: new Abstract: Vision-language-action (VLA) models improve robotic manipulation but remain vulnerable to compounding errors, sc…
Algorithm-Architecture Co-Design for Efficient VLA Inference via Speculative Inference and Verification
arXiv:2608.15636v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have demonstrated remarkable capabilities in the field of embodied AI, but t…
MistyPilot: Enabling Social-Robot Control through Multi-Agent LLM Skill Orchestration
arXiv:2608.15549v1 Announce Type: new Abstract: Programming small social robots from natural-language instructions requires more than invoking isolated APIs. In…
Temporal Logic Guided Universal Task Representations for Reinforcement Learning
arXiv:2608.15509v1 Announce Type: new Abstract: Task guided agents demonstrate strong performance in a wide range of complex tasks. However, most existing task …
Detachable Wire Drive : Reconfigurable Robot Architecture with Shared Actuators
arXiv:2608.15461v1 Announce Type: new Abstract: Reconfigurable robots offer significant potential for adapting to diverse tasks; however, conventional centraliz…
GUIDER: Evaluating Goal-Free Human Intent Inference for Teleoperated Manipulation on Real-Robot Data
arXiv:2608.15446v1 Announce Type: new Abstract: This paper presents an evaluation of a goal-free probabilistic framework for human intent inference during robot…
VTInstructor: Visual Trajectory Prompting for Navigation Instruction Generation in Continuous Environments
arXiv:2608.15284v1 Announce Type: new Abstract: Navigation instruction generation from ego-centric RGB video in continuous environments is an important yet chal…
Remember Smarter: Visual History Compressor and Hyperbolic Experience Space for Robotic Memory
arXiv:2608.15269v1 Announce Type: new Abstract: Long-horizon robot policies require compact access to recent observations and reusable experience without expand…
Low-Rank Dynamics-Effective Latent Carriers for Counterfactual Rollout in Learned World Models
arXiv:2608.15156v1 Announce Type: new Abstract: World models may predict the future without making clear which parts of their hidden state actually drive those …
StructRL: Structured Action-Space Exploration for Flow-Based VLAs
arXiv:2608.15139v1 Announce Type: new Abstract: Flow-based Vision-Language-Action (VLA) models are now widely used for continuous robotic manipulation, and onli…
From Continuous Design to Delay-Aware Discrete Synthesis: Guaranteed High-Bandwidth Joint Control for PMSM Drives
arXiv:2608.14937v1 Announce Type: new Abstract: The increasing dynamic demands of modern robotic joints require current controllers to achieve high bandwidth ov…
How Generalist uses human demonstration data for robot learning
Research around the Universal Manipulation Interface is enabling Generalist, Walden Robotics, and others to turn data into robot behaviors. The post How General…
Intermittent swimming promotes the energy efficiency of fish-like robot movements
Image credits: Xiangxiao Liu, Francois A. Longchamp, and Louis GeverBiorobotics Laboratory, EPFL Improving energy performance can effectively extend the time a …
Coverage Aware Active Evaluation for Failure Discovery with Paired Systems
arXiv:2608.13719v1 Announce Type: cross Abstract: Autonomous systems can fail in rare and heterogeneous ways, making real-world failure discovery difficult unde…
Simulation-Aware In-Context Policy Improvement for LLM-Aided Analog Layout Refinement
arXiv:2608.13767v1 Announce Type: cross Abstract: Analog IC layout design remains a labor-intensive iterative process dominated by simulation-driven refinement.…
MorphIt: Flexible Spherical Approximation of Robot Morphology for Representation-driven Adaptation
arXiv:2507.14061v3 Announce Type: replace Abstract: What if a robot could rethink its own morphological representation to better meet the demands of diverse tas…
AtomBridge: Agentic VLA Inference Plugin for Long-Horizon Tasks in Scientific Experiments
arXiv:2602.09430v2 Announce Type: replace Abstract: Robotic laboratories play a critical role in autonomous scientific discovery by enabling scalable, continuou…
CoViLLM: An Adaptive Human-Robot Collaborative Assembly Framework Using Large Language Models
arXiv:2603.11461v3 Announce Type: replace Abstract: With increasing demand for mass customization, traditional manufacturing robots that rely on rule-based oper…