News
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Latest in News
1860 storiesMitigating Covariate Shift in Imitation Learning for Autonomous Vehicles Using Latent Space Generative World Models
arXiv:2409.16663v5 Announce Type: replace Abstract: We propose the use of latent space generative world models to address the covariate shift problem in autonom…
Integrated Graph Search and Model Predictive Control for Smooth and Efficient Path Planning in Autonomous Vehicles
arXiv:2607.04259v1 Announce Type: new Abstract: Path planning is a fundamental component of autonomous vehicles, where achieving safe, comfortable, and dynamica…
OmniTacTune: Policy-Agnostic Real-World RL for Tactile Residual Adaptation of Visual Policies
arXiv:2607.03723v1 Announce Type: new Abstract: Visual policies learned from human videos, teleoperation, and robot demonstrations offer scalable motion priors,…
Hope for the Best, Prepare for the Worst: Occlusion-Aware Contingency Planning for Autonomous Vehicles
arXiv:2607.03155v1 Announce Type: new Abstract: The deployment of autonomous vehicles in urban environments introduces significant safety challenges, particular…
Scalable Dexterous Robot Learning with AR-based Remote Human-Robot Interactions
arXiv:2602.07341v2 Announce Type: replace-cross Abstract: This paper focuses on the scalable robot learning for manipulation in the dexterous robot arm-hand sys…
TiROD: Tiny Robotics Dataset and Benchmark for Continual Object Detection
arXiv:2409.16215v4 Announce Type: replace Abstract: Detecting objects with visual sensors is crucial for numerous mobile robotics applications, from autonomous …
InternVLA-A1.5: Unifying Understanding, Latent Foresight, and Action for Compositional Generalization
arXiv:2607.04988v1 Announce Type: new Abstract: Unified models for robot manipulation aim to equip one policy with both the semantic priors of pretrained VLMs a…
StageCraft: Execution Aware Mitigation of Distractor and Obstruction Failures in VLA Models
arXiv:2603.20659v2 Announce Type: replace Abstract: Large scale pre-training on text and image data along with diverse robot demonstrations has helped Vision La…
!Imperio, smolVLA: The Implications of Data Poisoning on Open Source Robotics
arXiv:2607.04146v1 Announce Type: new Abstract: This work establishes that trigger-word data poisoning of vision language action models is practical, while at t…
MOSAIC: Modular Scalable Autonomy for Intelligent Coordination of Heterogeneous Robotic Teams
arXiv:2601.23038v3 Announce Type: replace Abstract: Mobile robots have become indispensable for exploring hostile environments, such as in space or disaster rel…
GelNeuro: A Sensing-Computing Integrated Neuromorphic Tactile System for Texture Recognition
arXiv:2607.05241v1 Announce Type: new Abstract: Neuromorphic visuo-tactile sensing offers a promising paradigm for low-latency and low-power robotic perception.…
HALO-WA: Hybrid-Attention Latent-Guided Online Reinforcement Learning for World-Action Models
arXiv:2607.04265v1 Announce Type: new Abstract: World-action (WA) models can generate long-horizon action chunks for general-purpose robotic manipulation, but t…
Compressing the Validation Bottleneck: An Agentic Self-Driving Lab for Scientific Discovery
arXiv:2607.04508v1 Announce Type: cross Abstract: Agentic AI-for-Science can automate ideation, planning, and analysis, but final validation still depends on re…
$M^2$-VLA: Boosting Vision-Language Models for Generalizable Manipulation via Layer Mixture and Meta-Skills
arXiv:2604.24182v2 Announce Type: replace Abstract: Current Vision-Language-Action (VLA) models predominantly rely on end-to-end fine-tuning. While effective, t…
Green for Go, Red for No: Visual Grounding via Semantic Segmentation for VLA Navigation Policies
arXiv:2607.05122v1 Announce Type: cross Abstract: Vision-language-action (VLA) models enable robot navigation from natural language and visual goals, but remain…
Toward Personalized Social Robots for Child Well-being: Data Requirement Principles from a Recommender-System Perspective
arXiv:2607.05110v1 Announce Type: cross Abstract: Social robots are increasingly deployed in clinical settings to support the well-being of children, where effe…
Agent-driven Long-tail Simulation for Autonomous Driving
arXiv:2607.04331v1 Announce Type: new Abstract: Evaluating autonomous driving systems in closed-loop settings requires realistic and interactive simulation, yet…
Athena-WBC: Capability-Aligned Policy Experts for Long-Tail Humanoid Whole-Body Control
arXiv:2607.04837v1 Announce Type: new Abstract: Large-scale humanoid motion-tracking controllers are commonly improved by reallocating training effort: difficul…
HiMe: Hierarchical Embodied Memory for Long-Horizon Vision-Language-Action Control
arXiv:2607.03449v1 Announce Type: new Abstract: Current Vision-Language-Action (VLA) models excel at robotic manipulation but often struggle with non-Markovian …
AnchorVLA: Bridging Discrete Decisions and Continuous Trajectories for Vision-Language-Action Planning
arXiv:2607.03182v1 Announce Type: new Abstract: Autonomous driving planning requires translating navigation intent, traffic rules, dynamic interactions, and lan…
SurgAM: Surgical Affordance Map Prediction with Multimodal Feature Fusion for Robot Autonomy
arXiv:2607.04378v1 Announce Type: new Abstract: Surgical automation is being increasingly studied, yet bridging visual scene understanding with autonomous actio…
3D Cal: An Open-Source Software Library for Depth Reconstruction on Vision-Based Tactile Sensors
arXiv:2511.03078v3 Announce Type: replace Abstract: Tactile sensing plays a key role in enabling dexterous and reliable robotic manipulation, but realizing this…
TACO: TActile World Model as a Self-COrrector forScalable VLA Post-Training
arXiv:2607.02840v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have shown promising generalization in robotic manipulation, but they still …
GigaWorld-1: A Roadmap to Build World Models for Robot Policy Evaluation
arXiv:2607.02642v1 Announce Type: new Abstract: Evaluating embodied robot foundation models remains a critical bottleneck; unlike large language models efficien…
Learning 3D Affordances for Blade Insertion in Cluttered Stowing
arXiv:2607.02549v1 Announce Type: cross Abstract: Many manipulation tasks require reasoning about free-space affordances: discovering volumes where an extended …
WSA$_1$: a 3D-Centric World-Spatial-Action Model for Generalizable Robot Control
arXiv:2607.03941v1 Announce Type: new Abstract: Recent advances in embodied AI have established robot foundation models (RFMs) as the dominant approach for gene…
FLOAT Drone for Physical Interaction: Lateral Airflow Reduction, Wrench Modeling, and Adaptive Control
arXiv:2607.04260v1 Announce Type: new Abstract: Aerial physical interaction represents a promising direction for next-generation unmanned aerial vehicles (UAVs)…
Real-World Perturbation Testing of Autonomous Driving Systems
arXiv:2607.04953v1 Announce Type: cross Abstract: Autonomous Driving Systems (ADS) must operate reliably under diverse conditions, yet representative data for r…
KAM-WM: Kinematic Affordance Maps from Latent World Models for Robot Manipulation
arXiv:2607.04652v1 Announce Type: new Abstract: Learning manipulation from few demonstrations requires visual priors that capture not only where to interact, bu…
ObjRetarget: An Object-Aware Motion Retargeting Framework with Anthropomorphic Arm Constraints and Polyhedral Hand Modeling
arXiv:2607.03828v1 Announce Type: new Abstract: Learning robot dexterous manipulation from human manipulation videos requires reliably retargeting human intent …