Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
3542 storiesHand-centric Human-to-Robot Trajectory Transfer from Video Demonstrations via Open-World Contact Localization
arXiv:2606.10743v1 Announce Type: new Abstract: Learning from human video demonstrations remains challenging due to noisy hand-object interactions, unseen objec…
Pushing the Performance Limits in Autonomous Racing: Continuous Stability-Aware Adaptive Velocity Planning in Formula Student Driverless
arXiv:2606.10733v1 Announce Type: new Abstract: In autonomous racing, especially in competitions such as Formula Student Driverless, precise planning of the tar…
Vehicle Prediction Model for Enhanced MPC Path Tracking in Formula Student Driverless
arXiv:2606.10732v1 Announce Type: new Abstract: Autonomous race cars, such as in Formula Student Driverless, operate close to their physical handling limits. Th…
Self-Supervised Relevance Modelling in Autonomous Driving via Counterfactual Analysis
arXiv:2606.10688v1 Announce Type: new Abstract: Autonomous driving relies on computationally intensive perception pipelines to continuously detect and track obj…
UniDexTok: A Unified Dexterous Hand Tokenizer from Real Data
arXiv:2606.10683v1 Announce Type: new Abstract: Dexterous hands are essential for fine-grained manipulation, but their hardware designs vary substantially acros…
Planar-Sector LOS Guidance for Interception of Agile Targets with Lifting-Wing Quadcopters
arXiv:2606.10639v1 Announce Type: new Abstract: Autonomous visual interception of agile aerial targets is challenging due to unpredictable target motion, limite…
LieIPM: Lie Group Interior Point Method for Direct Trajectory Optimization of Rigid Bodies
arXiv:2606.10579v1 Announce Type: new Abstract: Designing dynamically feasible trajectories for rigid bodies is a fundamental problem in robotics. While direct …
VeriSpace: Spatially Grounded Action Verification for Vision-Language-Action Models
arXiv:2606.10568v1 Announce Type: new Abstract: Vision-language-action (VLA) models have shown strong promise for robotic manipulation, but their reliability at…
Uncovering Vulnerability of Vision-Language-Action Models under Joint-Level Physical Faults
arXiv:2606.10501v1 Announce Type: new Abstract: Deploying Vision-Language-Action (VLA) models in real robotic systems requires robustness not only to semantic a…
GuideWalk: Learning Unified Autonomous Navigation and Locomotion for Humanoid Robots across Versatile Terrains
arXiv:2606.10449v1 Announce Type: new Abstract: Humanoid robots have achieved strong locomotion capabilities, but reliable navigation on versatile terrains rema…
Information-Preserving Continuous Occupancy Mapping with Variance-Weighted Submap Joining
arXiv:2606.10442v1 Announce Type: new Abstract: Large-scale SLAM remains challenging due to accumulated trajectory drift and the increasing computational cost o…
UMI-Bench 1.0: An Open and Reproducible Real-World Benchmark for Tabletop Robotic Manipulation with UMI Data
arXiv:2606.10382v1 Announce Type: new Abstract: Real-robot evaluation is essential for understanding whether learned manipulation policies can operate reliably …
Test-time Adversarial Takeover: A Real-time Hijacking Interface against Robotic Diffusion Policies
arXiv:2606.10371v1 Announce Type: new Abstract: Diffusion-based action generation has become a foundational component of embodied AI, but its reliance on visual…
A Practical Recipe Towards Improving Sim-and-Real Correlation for VLA Evaluation
arXiv:2606.10366v1 Announce Type: new Abstract: Simulation has become an essential tool for evaluating and improving vision-language-action (VLA) policies, offe…
HiMem-WAM: Hierarchical Memory-Gated World Action Models for Robotic Manipulation
arXiv:2606.10363v1 Announce Type: new Abstract: World Action Models (WAMs) have emerged as a new powerful paradigm for embodied intelligence, learning action-re…
OMG: Omni-Modal Motion Generation for Generalist Humanoid Control
arXiv:2606.10340v1 Announce Type: new Abstract: Humanoid whole-body control has made significant progress in recent years, yet existing approaches remain limite…
SARM2: Multi-Task Stage Aware Reward Modeling for Self Improving Robotic Manipulation
arXiv:2606.10305v1 Announce Type: new Abstract: Fine-tuning vision-language-action (VLA) policies for long-horizon manipulation still relies heavily on behavior…
Improved Representation of Matrix Lie Group Operations through Tensor Notation
arXiv:2606.10289v1 Announce Type: new Abstract: Several recent papers have demonstrated the utility of using Lie groups within estimation problems, yielding imp…
MARCH: Model-Assisted Reinforcement Learning for the Perceptive Control of Humanoids over Sparse Footholds
arXiv:2606.10288v1 Announce Type: new Abstract: Perceptive bipedal locomotion over sparse terrain remains a difficult challenge: model-based methods are precise…
Hierarchical Policies from Verbal and Egocentric Human Signals for Natural Human-Robot Interaction
arXiv:2606.10276v1 Announce Type: new Abstract: For natural human-robot interaction, a robot must understand human intent expressed not only through language bu…
Locomotion analysis of a quadruped interacting with the lunar granular surface
arXiv:2606.10273v1 Announce Type: new Abstract: Deploying legged robots in extra-terrestrial environments includes many challenges due to complex terrain intera…
What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents
arXiv:2606.10267v1 Announce Type: new Abstract: Hierarchical vision-language-action (Hi-VLA) systems have emerged as a promising paradigm for complex robot mani…
YUBI: Yielding Universal Bidigital Interface for Bimanual Dexterous Manipulation at Scale
arXiv:2606.10244v1 Announce Type: new Abstract: We introduce Yielding Universal Bidigital Interface (YUBI), a finger-aligned gripper designed to enable intuitiv…
What Demonstration Curation Metrics Do to Your Policy
arXiv:2606.10229v1 Announce Type: new Abstract: We study whether demonstration-curation metrics that detect defective training episodes also improve the downstr…
Exploration of Foundation Model-Based Robots in Patient and Elderly Care
arXiv:2606.10208v1 Announce Type: new Abstract: Demand for older-adult and patient care is growing rapidly as populations age worldwide. Foundation models are i…
Flow Control: Steering Vision-Language-Action Models with Simple Real-Time Inputs
arXiv:2606.10180v1 Announce Type: new Abstract: We introduce flow control of vision-language-action (VLA) models, a simple and effective way to steer VLA action…
Efficient-WAM: A 1B-Parameter World-Action Model with Low-Cost Future Imagination
arXiv:2606.10040v1 Announce Type: new Abstract: World-Action Models (WAMs) have emerged as a promising paradigm for embodied control by coupling future visual p…
Robotic Nonprehensile Object Transportation with a Hanging Tray
arXiv:2606.10039v1 Announce Type: new Abstract: We consider the nonprehensile object transportation task known as the waiter's problem, in which a robot must mo…
GHOST: Hierarchical Sub-Goal Policies for Generalizing Robot Manipulation
arXiv:2606.10025v1 Announce Type: new Abstract: We present GHOST, a framework for learning visuomotor manipulation policies that generalize beyond the training …
Uncertainty-Aware Motion Planning for Autonomous Driving in Mixed Traffic Environment
arXiv:2606.09958v1 Announce Type: new Abstract: In mixed-traffic environments where autonomous and human-driven vehicles may co-exist, motion planning for auton…