News
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Latest in News
3285 storiesLook Where It Matters: Adaptive Visual Refinement for Vision-Language-Action Models
arXiv:2608.02197v1 Announce Type: new Abstract: Visual representations of VLA models remain unreliable for spatially precise robotic manipulation. We uncover th…
Local-Canonicalization Equivariant Graph Neural Networks for Sample-Efficient and Generalizable Swarm Robot Control
arXiv:2509.14431v2 Announce Type: replace Abstract: Multi-agent reinforcement learning (MARL) policies for swarm control often learn inefficiently and generaliz…
Spline Policy: A Structured Representation for Robot Policies
arXiv:2606.07386v2 Announce Type: replace Abstract: Modern imitation-learning policies for robot manipulation often represent actions as fixed-resolution action…
Latent-Centroid Steering: Single-Pass Classifier-Free Guidance for Command-Aligned Autonomous Driving
arXiv:2608.00237v1 Announce Type: cross Abstract: Vision-language models (VLMs) have recently emerged as a promising paradigm for end-to-end autonomous driving,…
AffordTrajDP: Dynamic Affordance-Guided Visuomotor Policy Learning for Robotic Manipulation
arXiv:2608.01603v1 Announce Type: new Abstract: Affordance-guided imitation learning has shown impressive performance in robotic manipulation tasks by compressi…
PRISM: Privileged Probabilistic Latent Supervision for End-to-End Autonomous Driving Motion Planning
arXiv:2608.01201v1 Announce Type: new Abstract: End-to-end autonomous driving (E2E AD) systems integrate perception, prediction, and planning into a single diff…
ORCESTRA: VLM-driven Visual Robot programming in Mixed Reality
arXiv:2608.00775v1 Announce Type: new Abstract: ORCESTRA is a mixed-reality system for programming robot digital twins through no-code waypoint teaching and lan…
HapticVLA: Contact-Rich Manipulation via Vision-Language-Action Model without Inference-Time Tactile Sensing
arXiv:2603.15257v2 Announce Type: replace Abstract: Tactile sensing is a crucial capability for Vision-Language-Action (VLA) architectures, as it enables dexter…
Hybrid Attention Estimation Pipeline for Adaptive HRI Using an Expressive Robotic Head
arXiv:2608.00284v1 Announce Type: new Abstract: This paper presents an applied case study on hybrid visual attention estimation for human-robot interaction usin…
DriveCode: Domain Specific Numerical Encoding for LLM-Based Autonomous Driving
arXiv:2603.00919v3 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown great promise for autonomous driving. However, discretizing nu…
DiffPhysCam: Differentiable Physics-Based Camera Simulation for Inverse Rendering and Embodied AI
arXiv:2508.08831v2 Announce Type: replace-cross Abstract: Generating synthetic images that closely mimic those from real cameras is instrumental in training vis…
Faster-WAM: Do World Action Models Need Deep Action Modules?
arXiv:2608.02365v1 Announce Type: cross Abstract: World Action Models (WAMs) couple robot action prediction with video world models. Existing WAMs with shared-b…
CoWAM: Coordination Contracts for Selective Policy Intervention with WAMs
arXiv:2608.02578v1 Announce Type: new Abstract: World Action Models (WAMs) augment robot policies with action-conditioned predicted futures, but a plausible fut…
Teleopit: A Full-Embodiment Humanoid Teleoperation System
arXiv:2608.01834v1 Announce Type: new Abstract: Humanoid teleoperation for demonstration collection requires coordinated whole-body motion, continuous dexterous…
WAM-Diff2: Hierarchical AR-to-Diffusion Distillation for Highly Efficient Autonomous Driving VLA
arXiv:2608.01035v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have emerged as a prominent paradigm for end-to-end autonomous driving; howe…
When Prompts Control Robots: Prompt Injection Attacks in Multi-Agent Robotic Systems
arXiv:2608.00747v1 Announce Type: new Abstract: Large language models are increasingly integrated into autonomous robotic systems for task planning and control,…
SelfWAM: A Self-Grounded Unified World Action Model for Fast Robot Control
arXiv:2608.00725v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling actions and future observations. Ho…
EndoWAM: A Grounded World-Action Model for Generalizable Endoscopic Navigation
arXiv:2608.01221v1 Announce Type: new Abstract: Autonomous endoscopic navigation can reduce clinicians' operational burden, yet robust control remains challengi…
Adaptive Human-Robot Collaborative Painting Combining Preference-Based Optimization and Dynamic Motion Primitives
arXiv:2608.01981v1 Announce Type: new Abstract: This work presents a human-centered collaborative framework that integrates Preference-Based Optimization (PBO) …
Rapid Embodiment Adaptation for Quadrupedal Locomotion
arXiv:2608.01506v1 Announce Type: new Abstract: Humans readily adapt their movements as their bodies change through aging, injury, or load carrying, but learnin…
Situation Aware Frontier Prioritization for Quadruped Search and Rescue
arXiv:2608.02571v1 Announce Type: new Abstract: Quadruped robots are a promising platform for search and rescue missions because they can navigate cluttered ind…
Identifying and Exploiting Structure in Robot Co-Design
arXiv:2604.11768v2 Announce Type: replace Abstract: Co-design of a robot's morphology and control is a high-dimensional search problem. Efficient search depends…
QuASH: Using Natural-Language Heuristics to Query Visual-Language Robotic Maps
arXiv:2510.14546v2 Announce Type: replace Abstract: Embeddings from Visual-Language Models are increasingly utilized to represent semantics in robotic maps, off…
Weights or Skills? A Survey of Robot-Learning Techniques: from Action-Predicting Weights to Robots that Write their Own Skills
arXiv:2608.01851v1 Announce Type: new Abstract: Robot learning is splitting into two bets: policies that bake competence into frozen weights (vision-language-ac…
A drone that learns to efficiently find non-uniformly distributed objects in agricultural fields: from simulation to the real world
arXiv:2505.09278v2 Announce Type: replace Abstract: Drones are promising for data collection in precision agriculture but are limited by battery capacity. Drone…
Learning-Based Motion Planning for Dynamic Environments: From Foundational Algorithms to Emerging Paradigms
arXiv:2608.00625v1 Announce Type: new Abstract: Motion planning in dynamic environments is a fundamental problem in robotics, aiming to generate safe and effici…
TravKAN: Fast and Interpretable Nonlinear Traversability Analysis with Kolmogorov-Arnold Networks
arXiv:2608.02320v1 Announce Type: new Abstract: Traversability analysis is a fundamental capability for autonomous mobile robots operating in unstructured envir…
Who Is Responsible? Self-Adaptation Under Multiple Concurrent Uncertainties With Unknown Sources in Complex ROS-Based Systems
arXiv:2504.20477v4 Announce Type: replace Abstract: Robotic systems increasingly operate in dynamic, unpredictable environments, where tightly coupled sensors a…
Bicycle Acrobatics with Reinforcement Learning
arXiv:2608.00880v1 Announce Type: new Abstract: Bicycle robots are fast and energy efficient, but their simple mechanical design and their underactuated and non…
GeminiPainter's sequence-formed pipeline comprised of perception, cognition, planning, and action stages
arXiv:2608.00829v1 Announce Type: new Abstract: We present an autonomous robotic portrait-generation system combining real-time face detection, AI-based sketch …