News
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Latest in News
3285 storiesRobots as Tokens: Unified Diffusion Transformer for Coordinated Multi-Robot Trajectory Generation
arXiv:2606.15550v1 Announce Type: new Abstract: The success of generative models in language and visual generation has inspired extensive applications to genera…
ControlMap: Controllable High-Definition Map Generation for Traffic Scenario Simulation
arXiv:2606.15930v1 Announce Type: new Abstract: Simulation is central to validating autonomous driving systems, yet current pipelines are limited by insufficien…
RHO: Your Coding Agent is Secretly a Roboticist
arXiv:2606.16458v1 Announce Type: new Abstract: Code-as-Policies (CaP) has shown that large language models (LLMs) can write code to solve robotics tasks by com…
Robots that Collaborate: Sequential Asymmetric Imitation for Learning Coupled Robot Policies
arXiv:2606.16490v1 Announce Type: new Abstract: Collaborative mobile manipulation requires robots to coordinate with a partially observed partner while physical…
V2P-Manip: Learning Dexterous Manipulation from Monocular Human Videos
arXiv:2606.16436v1 Announce Type: new Abstract: Achieving autonomous robotic dexterous manipulation requires precise, human-like action sequences at scale. As a…
PATCH: Action-Chunk-Conditioned Latent Patch Innovation Monitoring for Robot Manipulation
arXiv:2606.16690v1 Announce Type: new Abstract: Learning-based manipulation policies have made substantial progress in real-world robot manipulation, particular…
ATOM-Bench: A Real-World Benchmark for Atomic Skills and Compositional Generalization in Manipulation Policies
arXiv:2606.16826v1 Announce Type: new Abstract: Generalist manipulation policies are increasingly presented as foundation models for robotic control, but their …
LOPAL: Local Performance-Aware Active Learning from Imperfect Demonstrations
arXiv:2606.16888v1 Announce Type: new Abstract: Learning from Demonstration (LfD) enables intuitive robot skill acquisition by allowing robots to learn directly…
MotionVLA: Vision-Language-Action Model for Humanoid Motion
arXiv:2606.15142v1 Announce Type: cross Abstract: Generating realistic humanoid motion from scene images and text involves both low-frequency pose semantics and…
X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining
arXiv:2606.14752v1 Announce Type: cross Abstract: Modern Vision-Language-Action (VLA) models must bridge pretrained vision-language reasoning and precise contin…
T-Rex: Tactile-Reactive Dexterous Manipulation
arXiv:2606.17055v1 Announce Type: new Abstract: The ability to react dynamically to tactile signals has long been considered crucial to agile human-level dexter…
Beyond English: Uncovering the Multilingual Gap in Vision-Language-Action Models
arXiv:2606.15714v1 Announce Type: cross Abstract: Vision-Language-Action models have recently demonstrated promising capabilities in learning generalist robot p…
Towards mm-Level Accurate UWB Radar: High-Accuracy Phase-Based Obstacle Detection through Multi-Channel Fusion
arXiv:2606.16657v1 Announce Type: cross Abstract: Accurate, tag-free distance estimation with ultrawideband (UWB) radar is essential for applications such as au…
Intelligent Sailing Model for Open Sea Navigation
arXiv:2501.04988v2 Announce Type: replace Abstract: Autonomous vessels potentially enhance safety and reliability of seaborne trade. To facilitate the developme…
DemoDiffusion: One-Shot Human Imitation using pre-trained Diffusion Policy
arXiv:2506.20668v3 Announce Type: replace Abstract: We propose DemoDiffusion, a simple method for enabling robots to perform manipulation tasks by imitating a s…
Latent Action Pretraining Through World Modeling
arXiv:2509.18428v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have gained popularity for learning robotic manipulation tasks that foll…
CLAP: Contrastive Latent Action Pretraining for Learning Vision-Language-Action Models from Human Videos
arXiv:2601.04061v2 Announce Type: replace Abstract: Generalist Vision-Language-Action models remain constrained by the scarcity of robotic data relative to the …
LaST$_{0}$: Latent Spatio-Temporal Chain-of-Thought for Robotic Vision-Language-Action Model
arXiv:2601.05248v4 Announce Type: replace Abstract: Vision-Language-Action (VLA) models have recently shown strong generalization, with some approaches seeking …
Neural Minimum-Distance Estimation for Collision-Aware Operation of Multi-Arm Laparoscopy Surgical Robots Through Learning-from-Simulation
arXiv:2601.15459v2 Announce Type: replace Abstract: This study presents an integrated framework for enhancing the safety and operational efficiency of robotic a…
IVRA: Improving Visual-Token Relations for Robot Action Policy with Training-Free Hint-Based Guidance
arXiv:2601.16207v2 Announce Type: replace Abstract: Many Vision-Language-Action (VLA) models flatten image patches into a 1D token sequence, weakening the 2D sp…
Bimanual High-Density EMG Control for In-Home Mobile Manipulation by Users with Quadriplegia
arXiv:2602.02773v2 Announce Type: replace Abstract: Mobile manipulators in the home can enable people with cervical spinal cord injury (cSCI) to perform daily p…
HiCrowd: Hierarchical Crowd Flow Alignment for Dense Human Environments
arXiv:2602.05608v3 Announce Type: replace Abstract: Navigating through dense human crowds remains a significant challenge for mobile robots. A key issue is the …
FineVLA: Fine-Grained Instruction Alignment for Steerable Vision-Language-Action Policies
arXiv:2605.27284v2 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are increasingly expected to not only complete robot tasks, but also fol…
Seeing Roads Through Words: A Language-Guided Framework for RGB-T Driving Scene Segmentation
arXiv:2602.07343v2 Announce Type: replace-cross Abstract: Robust semantic segmentation of road scenes under adverse illumination, lighting, and shadow condition…
Systematic Evaluation of Novel View Synthesis for Video Place Recognition
arXiv:2603.05876v2 Announce Type: replace-cross Abstract: The generation of synthetic novel views has the potential to positively impact robot navigation in sev…
Artists' Views on Robotics Involvement in Painting Productions
arXiv:2510.07063v3 Announce Type: replace-cross Abstract: As robotic technologies evolve, their potential in artistic creation becomes an increasingly relevant …
Bio-inspired decision making in robot swarms under biases
arXiv:2509.07561v2 Announce Type: replace-cross Abstract: Minimalistic robot swarms offer a scalable, robust, and cost-effective approach to performing complex …
DynNPC: Finding More Violations Induced by ADS in Simulation Testing through Dynamic NPC Behavior Generation
arXiv:2411.19567v3 Announce Type: replace-cross Abstract: Recently, a number of simulation testing approaches have been proposed to generate diverse driving sce…
Imitating What Works: Simulation-Filtered Modular Policy Learning from Human Videos
arXiv:2602.13197v2 Announce Type: replace Abstract: The ability to learn manipulation skills by watching videos of humans has the potential to unlock a new sour…
Simplifying ROS2 controllers with a modular architecture for robot-agnostic reference generation
arXiv:2601.08514v2 Announce Type: replace Abstract: This paper introduces a novel modular architecture for ROS2 that decouples the logic required to acquire, va…