News
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Latest in News
3285 storiesSkyDrive: Learning to Drive in a New City from Aerial Traffic Monitoring
arXiv:2608.25142v1 Announce Type: new Abstract: Autonomous driving has made remarkable progress through imitation learning with massive human demonstration data…
Latent Chain-of-Thought World Modeling for End-to-End Driving
arXiv:2512.10226v3 Announce Type: replace-cross Abstract: Recent Vision-Language-Action (VLA) models for autonomous driving explore inference-time reasoning as …
Minimalist Visual Inertial Odometry
arXiv:2605.19990v2 Announce Type: replace Abstract: Visual-Inertial Odometry (VIO), which is critical to mobile robot navigation, uses cameras with a large numb…
Scene2Demo: Self-Evolving Embodied Data Generation via Object-Action Graph
arXiv:2602.12065v2 Announce Type: replace Abstract: We present Scene2Demo, a self-evolving framework for offline embodied data generation. Given a single real-w…
Zero-WAM: In-Context World-Action Modeling from Human Videos for Open-Ended Task Generalization
arXiv:2608.26103v1 Announce Type: new Abstract: Zero-shot cross-task generalization, where a policy must execute manipulation tasks never seen during training, …
Fast Generative Grasping via Lie Group-Constrained MeanFlow
arXiv:2608.26076v1 Announce Type: new Abstract: Grasp synthesis is a core task in robotic manipulation, for which the solution typically forms a multimodal dist…
A Statistical Audit of Physical AI Benchmark Redundancy
arXiv:2608.25940v1 Announce Type: new Abstract: Physical AI models are evaluated on suites of benchmarks that differ across model reports, leaving the model-by-…
LM-X: Explainable Action Modeling with Progress, Event, and Uncertainty Prediction for Generalist Robot Manipulation
arXiv:2608.25757v1 Announce Type: new Abstract: Generalist vision--language--action (VLA) policies learn long-horizon behavior mainly through short-horizon acti…
GaussianDream++: Efficient 3D Gaussian World Modeling for Robotic Manipulation
arXiv:2608.25659v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies have advanced language-conditioned robotic manipulation, yet action-imitat…
RA-VLA: Retrieval-Augmented VLA for Test-Time Adaptation
arXiv:2608.25585v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models provide a versatile foundation for general robotic manipulation, yet they ex…
A Taxonomy of Construction Task Activities for Robot Workers
arXiv:2608.25395v1 Announce Type: new Abstract: Recent vision-language-action models offer a path toward robots with broader repertoires than conventional task-…
RAEM: Robust Autonomous Exploration for Multi-Floor Environments with a Quadruped Robot
arXiv:2608.25366v1 Announce Type: new Abstract: In this paper, we propose RAEM, a robust autonomous exploration framework for quadruped robots operating in mult…
ROS2 Connect: A new ROS2 over WAN Solution
arXiv:2608.25102v1 Announce Type: new Abstract: The Robot Operating System 2 (ROS2) has become a widely adopted framework for the development of distributed rob…
IDS Imaging adds Nion ToF sensor to its portfolio of 3D cameras
IDS said Nion is suitable for applications where precise 3D data is required at real-world process speeds, such as logistics and robotics. The post IDS Imaging …
Gatik brings in $200M to continue expanding autonomous trucking operations
Gatik said the funding will help it expand its model built around high-frequency regional routes connecting distribution centers and stores. The post Gatik brin…
Bedrock Robotics’ first operator-free excavator deployments take off
Bedrock Robotics deploys autonomous excavators to active job sites, tackling construction labor shortages with AI-powered heavy machinery. The post Bedrock Robo…
How green is your robot? And other awkward questions
Deborah Lupton / Servers in a Landscape / Licenced by CC-BY 4.0 By Emmet Cole Robots clean rivers and sort waste, monitor ecosystems, and inspect renewable-ener…
CARE: Camera-Residual Reserves for First Sightings in Adaptive LiDAR Sensing
arXiv:2608.24282v1 Announce Type: cross Abstract: Adaptive LiDAR scanning concentrates a limited sensing budget on regions of interest predicted from past objec…
VGGT-DP: Generalizable Robot Control via Vision Foundation Models
arXiv:2509.18778v2 Announce Type: replace Abstract: Visual imitation learning frameworks allow robots to learn manipulation skills from expert demonstrations. W…
A Robust Task-Level Control Architecture for Learned Dynamical Systems
arXiv:2511.09790v2 Announce Type: replace Abstract: Dynamical system (DS)-based learning from demonstration (LfD) is a powerful tool for generating motion plans…
From Dialogue to Execution: Mixture-of-Agents Assisted Interactive Planning for Behavior Tree-Based Long-Horizon Robot Execution
arXiv:2603.01113v2 Announce Type: replace Abstract: Interactive task planning with large language models (LLMs) lets robots generate high-level action plans fro…
AURASeg: Attention-Guided Upsampling with Residual-Assisted Boundary Refinement for Drivable-Area Segmentation
arXiv:2510.21536v5 Announce Type: replace Abstract: Free-space segmentation is essential for autonomous robots to identify drivable regions and navigate safely …
Do Robotic World Models Really Follow Actions? Diagnosing and Aligning Action-Conditioned Generation for Policy Learning
arXiv:2608.24885v1 Announce Type: new Abstract: Action-conditioned world models are increasingly used as learned simulators for policy evaluation and improvemen…
Latent Action as Intention Enables Efficient Future Imagination for World Action Models
arXiv:2608.24882v1 Announce Type: new Abstract: World action models (WAMs) improve robot control by modeling how observations evolve, but generating future obse…
One-Shot Learning from Demonstration of Contact-Rich Robotic Manipulation by Identifying Physical Interactions
arXiv:2608.24741v1 Announce Type: new Abstract: Learning from Demonstration (LfD) allows robots to learn manipulation tasks directly from humans, thereby suppor…
NVIDIA Cosmos-H-Dreams: Real-Time Generative Physics Simulation for Surgical Robotics
arXiv:2608.24199v1 Announce Type: new Abstract: Generative simulation for surgical robotics still lacks real-time interaction. Physical-robot experiments, often…
Coverage Planning for Robotic Tooth Preparation in Densely Constrained Environments
arXiv:2608.24155v1 Announce Type: new Abstract: Tooth preparation refers to the controlled removal of tooth structure to create an optimal substrate for fixed r…
PonderPounce: A Pretrained MLLM as an Episode Context Engine for Robot Control
arXiv:2608.24115v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) can integrate long visual histories, reason under partial observability…
Hierarchical Skill Retrieval for Data-Efficient Adaptation of Vision-Language-Action Models
arXiv:2608.24042v1 Announce Type: new Abstract: While Vision-Language-Action (VLA) models pretrained on large-scale robot datasets provide a strong foundation f…
NeurRAFT: Robot Motion Planning via Anchor-Level Flow Matching with Clearance-Aware Preference Tuning
arXiv:2608.24026v1 Announce Type: new Abstract: Recent end-to-end neural motion planners generate trajectories from raw sensor observations, avoiding the privil…