News
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Latest in News
1860 storiesAssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly
arXiv:2604.08983v2 Announce Type: replace Abstract: Spatial reasoning is a fundamental capability for embodied intelligence, especially for fine-grained manipul…
From Digital to Physical: Digital Agents as Autonomous Coaches for Physical Intelligence
arXiv:2601.21570v2 Announce Type: replace-cross Abstract: The field of Embodied AI is witnessing a rapid evolution toward general-purpose robotic systems, fuele…
Action-Effect Memory Pretraining for Robot Manipulation
arXiv:2606.12499v1 Announce Type: new Abstract: We present AEM, an Action-Effect Memory pretraining framework for robot manipulation that learns compact tempora…
EgoMoD: Predicting Global Maps of Dynamics from Local Egocentric Observations
arXiv:2603.00167v2 Announce Type: replace Abstract: Efficient navigation in dynamic environments requires anticipating how motion patterns evolve beyond the rob…
Miniature Testbed for Validating Multi-Agent Cooperative Autonomous Driving
arXiv:2511.11022v2 Announce Type: replace Abstract: Cooperative autonomous driving, which extends vehicle autonomy by enabling real-time collaboration between v…
GLIDE: A Coordinated Aerial-Ground Framework for Search and Rescue in Unknown Environments
arXiv:2509.14210v4 Announce Type: replace Abstract: We present a cooperative aerial-ground search-and-rescue (SAR) framework that pairs two unmanned aerial vehi…
Data-Driven Soft Robot Control via Adiabatic Spectral Submanifolds
arXiv:2503.10919v3 Announce Type: replace Abstract: The mechanical complexity of soft robots creates significant challenges for their model-based control. Speci…
GAE: Unleashing Physical Potential of VLM with Generalizable Action Expert
arXiv:2510.03896v2 Announce Type: replace-cross Abstract: Vision-language models demonstrate strong reasoning and planning abilities, yet grounding these predic…
Safety Case Patterns for VLA-based driving systems: Insights from SimLingo
arXiv:2603.16013v3 Announce Type: replace Abstract: Vision-Language-Action (VLA)-based driving systems represent a significant paradigm shift in autonomous driv…
Lyapunov-Based PI-Like Control for Robust Trajectory Tracking of a Four-Wheel Independently Driven and Steered Robot: Design and Experimental Validation
arXiv:2602.15424v2 Announce Type: replace Abstract: In this paper, a Lyapunov-based synthesis of a PI-like controller is proposed for robust trajectory tracking…
Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video
arXiv:2511.18322v4 Announce Type: replace Abstract: Learning soft continuum robot (SCR) dynamics from video offers flexibility but existing methods lack interpr…
Adaptive Model-Predictive Control of a Soft Continuum Robot Using a Physics-Informed Neural Network Based on Cosserat Rod Theory
arXiv:2508.12681v3 Announce Type: replace Abstract: Dynamic control of soft continuum robots (SCRs) holds great potential for expanding their applications, but …
Heterogeneous LiDAR Early Fusion and Learned Re-Ranking Strategy for Robust Long-Term Place Recognition in Unstructured Environments
arXiv:2606.13503v1 Announce Type: cross Abstract: Robust localization in unstructured environments, such as agricultural fields, is a critical challenge for aut…
Diffusion Transformer World-Action Model for AV Scene Prediction
arXiv:2606.12987v1 Announce Type: cross Abstract: Action-conditioned world models let an autonomous vehicle predict future camera scenes from its own planned co…
SemanticXR: Low Power and Real-time Queryable Semantic Mapping with an Object-Level Device-Cloud Architecture
arXiv:2606.12849v1 Announce Type: cross Abstract: Semantic mapping is a core service that enables grounded interactions in emerging Extended Reality (XR) applic…
Improving Robotic Generalist Policies via Flow Reversal Steering
arXiv:2606.13675v1 Announce Type: new Abstract: Generalist policies can learn a wide range of skills from diverse robot datasets. In order to solve or improve o…
SPARC: Reliable Spatial Annotations from Robot Demonstrations at Scale
arXiv:2606.13497v1 Announce Type: new Abstract: This work introduces Spatial Annotations from Robot Demonstrations with Reliability Calibration (SPARC), a risk-…
GIVE: Grounding Human Gestures in Vision-Language-Action Models
arXiv:2606.13435v1 Announce Type: new Abstract: Human communication is inherently multimodal, where language is often accompanied by non-verbal cues such as ges…
EMG-Based Adaptation of Anisotropic Virtual Fixtures for Robot-Assisted Surgical Resection and Dissection
arXiv:2606.13340v1 Announce Type: new Abstract: In this paper, we address the development of an adaptive assistance system for robot-assisted laparoscopic surge…
See Selectively, Act Adaptively: Dual-Level Structural Decomposition for Bimanual Robot Manipulation
arXiv:2606.13279v1 Announce Type: new Abstract: In bimanual robotic manipulation, task-relevant visual information varies with the task stage and context, while…
Proprioceptive-visual correspondence enables self-other distinction in humanoid robots
arXiv:2606.13222v1 Announce Type: new Abstract: Distinguishing self from others is a prerequisite for social intelligence, yet humanoid robots that increasingly…
FTP-1: A Generalist Foundation Tactile Policy Across Tactile Sensors for Contact-Rich Manipulation
arXiv:2606.13102v1 Announce Type: new Abstract: Despite the success of vision-based generalist robotic policies, existing tactile-based policies remain tied to …
RoboProcessBench: Benchmarking Process-Aware Understanding in Vision-Language Robotic Manipulation
arXiv:2606.13040v1 Announce Type: new Abstract: Vision-language models (VLMs) are increasingly explored as visual critics, reward generators, and failure detect…
GenHOI: Contact-Aware Humanoid-Object Interaction by Imitating Generated Videos without Task-Specific Training
arXiv:2606.12995v1 Announce Type: new Abstract: Humanoid-Object Interaction (HOI) is a fundamental capability for humanoid robots, yet it remains challenging du…
Trajectory-Level Redirection Attacks on Vision-Language-Action Models
arXiv:2606.12978v1 Announce Type: new Abstract: Vision-language-action (VLA) policies bring natural language into closed-loop robot control, enabling robots to …
Towards Reliable Sequential Object Picking in Clutter: The Runner-up Solution to RGMC 2025
arXiv:2606.12954v1 Announce Type: new Abstract: As a long-standing challenge in robotic manipulation, stable and efficient grasping in cluttered environments is…
An Embodied Simulation Platform, Benchmark, and Data-Efficient Augmentation Framework for Wet-Lab Robotics
arXiv:2606.12936v1 Announce Type: new Abstract: Wet-lab robots can improve the reproducibility, throughput, and safety of biomedical experiments, but scaling th…
DARRMS -- An Efficient Algorithm for Dynamic Attention Radius in Resource-Constrained Multi-Agent Systems
arXiv:2606.12614v1 Announce Type: new Abstract: Multi-agent systems are integral tools for various domains such as robotics, cybersecurity, and autonomous vehic…
From Imitation to Alignment: Human-Preference Flow Policies for Long-Horizon Sidewalk Navigation
arXiv:2606.12603v1 Announce Type: new Abstract: Autonomous long-horizon sidewalk navigation is essential for micro-mobility applications such as robotic food de…
Trojan Attacks on Neural Network Controllers for Robotic Systems
arXiv:2602.05121v2 Announce Type: replace-cross Abstract: Neural network controllers are increasingly deployed in robotic systems for tasks such as trajectory t…