News
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Latest in News
1860 storiesLatent Action Pretraining Through World Modeling
arXiv:2509.18428v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have gained popularity for learning robotic manipulation tasks that foll…
CLAP: Contrastive Latent Action Pretraining for Learning Vision-Language-Action Models from Human Videos
arXiv:2601.04061v2 Announce Type: replace Abstract: Generalist Vision-Language-Action models remain constrained by the scarcity of robotic data relative to the …
LaST$_{0}$: Latent Spatio-Temporal Chain-of-Thought for Robotic Vision-Language-Action Model
arXiv:2601.05248v4 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently shown strong generalization, with some approaches seeking …
Neural Minimum-Distance Estimation for Collision-Aware Operation of Multi-Arm Laparoscopy Surgical Robots Through Learning-from-Simulation
arXiv:2601.15459v2 Announce Type: replace Abstract: This study presents an integrated framework for enhancing the safety and operational efficiency of robotic a…
IVRA: Improving Visual-Token Relations for Robot Action Policy with Training-Free Hint-Based Guidance
arXiv:2601.16207v2 Announce Type: replace Abstract: Many Vision-Language-Action (VLA) models flatten image patches into a 1D token sequence, weakening the 2D sp…
Bimanual High-Density EMG Control for In-Home Mobile Manipulation by Users with Quadriplegia
arXiv:2602.02773v2 Announce Type: replace Abstract: Mobile manipulators in the home can enable people with cervical spinal cord injury (cSCI) to perform daily p…
HiCrowd: Hierarchical Crowd Flow Alignment for Dense Human Environments
arXiv:2602.05608v3 Announce Type: replace Abstract: Navigating through dense human crowds remains a significant challenge for mobile robots. A key issue is the …
FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies
arXiv:2605.27284v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are increasingly expected to not only complete robot tasks, but also fol…
Seeing Roads Through Words: A Language-Guided Framework for RGB-T Driving Scene Segmentation
arXiv:2602.07343v2 Announce Type: replace-cross Abstract: Robust semantic segmentation of road scenes under adverse illumination, lighting, and shadow condition…
Systematic Evaluation of Novel View Synthesis for Video Place Recognition
arXiv:2603.05876v2 Announce Type: replace-cross Abstract: The generation of synthetic novel views has the potential to positively impact robot navigation in sev…
Artists' Views on Robotics Involvement in Painting Productions
arXiv:2510.07063v3 Announce Type: replace-cross Abstract: As robotic technologies evolve, their potential in artistic creation becomes an increasingly relevant …
Bio-inspired decision making in robot swarms under biases
arXiv:2509.07561v2 Announce Type: replace-cross Abstract: Minimalistic robot swarms offer a scalable, robust, and cost-effective approach to performing complex …
DynNPC: Finding More Violations Induced by ADS in Simulation Testing through Dynamic NPC Behavior Generation
arXiv:2411.19567v3 Announce Type: replace-cross Abstract: Recently, a number of simulation testing approaches have been proposed to generate diverse driving sce…
Imitating What Works: Simulation-Filtered Modular Policy Learning from Human Videos
arXiv:2602.13197v2 Announce Type: replace Abstract: The ability to learn manipulation skills by watching videos of humans has the potential to unlock a new sour…
Simplifying ROS2 controllers with a modular architecture for robot-agnostic reference generation
arXiv:2601.08514v2 Announce Type: replace Abstract: This paper introduces a novel modular architecture for ROS2 that decouples the logic required to acquire, va…
Multi-Robot Motion Planning from Vision and Language using Heat-Inspired Diffusion
arXiv:2512.13090v2 Announce Type: replace Abstract: Diffusion models have recently emerged as powerful tools for robot motion planning by capturing the multi-mo…
Bayesian Optimization for Learning Nonlinear MPC in Autonomous Agent Navigation
arXiv:2606.14763v1 Announce Type: new Abstract: Real-time autonomous navigation in dynamic, unknown environments remains a fundamental challenge for mobile robo…
OmniVTLA: Vision-Tactile-Language-Action Models with Semantic-Aligned Tactile Sensing
arXiv:2508.08706v3 Announce Type: replace Abstract: Recent vision-language-action (VLA) models build upon vision-language foundations, and have achieved promisi…
MVOFormer: Flow-Semantic Transformer for Robust Monocular Visual Odometry
arXiv:2606.16474v1 Announce Type: cross Abstract: Monocular visual odometry (MVO) is foundational to autonomous navigation and robotic localization. However, ex…
Towards Next-Generation Healthcare: A Survey of Medical Embodied AI for Perception, Decision-Making, and Action
arXiv:2606.15647v1 Announce Type: cross Abstract: Foundation models have demonstrated impressive performance in enhancing healthcare efficiency across a wide ra…
VLALeaks: Membership Inference Attacks against Vision-Language-Action Models
arXiv:2606.15165v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models enable end-to-end robot control and have garnered widespread attention. Ho…
Geometric Action Model for Robot Policy Learning
arXiv:2606.17046v1 Announce Type: new Abstract: Generalist robot policies must follow user instructions while reasoning about how objects, cameras, and robot ac…
R2RDreamer: 3D-aware Data Augmentation for Spatially-generalized 2D Manipulation Policies
arXiv:2606.17040v1 Announce Type: new Abstract: Spatial generalization is critical for imitation-learned manipulation policies, but achieving it typically requi…
ROVE: Unlocking Human Interventions for Humanoid Manipulation via Reinforcement Learning
arXiv:2606.17011v1 Announce Type: new Abstract: Human interventions provide crucial corrective signals for post-training Vision-Language-Action (VLA) models. Ho…
When Should a Robot Replan? Regret-Guided Update Scheduling in Time-Varying MDPs
arXiv:2606.16972v1 Announce Type: new Abstract: Robots operating in non-stationary environments must continually adapt their policies as the dynamics drift, but…
Pride and Prejudice: Toward an Information-Theoretic Framework for Mutually Communicative Driver Behavior Modeling
arXiv:2606.16735v1 Announce Type: new Abstract: Mixed autonomy driving becomes unsafe and inefficient when autonomous vehicles (AVs) and human-driven vehicles (…
VENOM: Versatile Embodied Network for Omni-bodied Motion tracking
arXiv:2606.16696v1 Announce Type: new Abstract: Achieving expert-level expressive full-body motion tracking across multiple humanoids solely from demonstration …
WaveSync: Constrained Wavefront Optimization for Synchronized Co-Speech Gestures in Humanoid Robots
arXiv:2606.16600v1 Announce Type: new Abstract: Expressive co-speech gestures are crucial for natural human-robot interaction, but generating them on physical h…
Steering Generative Reinforcement Learning into Stable Robotic Controller
arXiv:2606.16572v1 Announce Type: new Abstract: Diffusion and flow-based generative policies provide a powerful policy class for reinforcement learning by induc…
ADAPT: Analytical Disturbance-Aware Policy Training for Humanoid Locomotion
arXiv:2606.16542v1 Announce Type: new Abstract: Humanoids deployed in human-centered environments must handle force-interactive tasks, where external contacts i…