Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
6273 storiesRisk-Aware Decision-Making for Autonomous Overtaking: A World Model-Based Mixture-of-Experts Framework
arXiv:2609.00385v1 Announce Type: new Abstract: Autonomous highway overtaking demands foresighted decision-making to handle complex interactions, stochastic tra…
Visko launches Orbis live model and closes pre-seed funding round
Visko says Orbis streams 4K video at 24 frames per second, responds to user intervention, and sustains hour-scale generation without drift. The post Visko launc…
Scale-Plan: Scalable Language-Enabled Task Planning for Heterogeneous Multi-Robot Teams
arXiv:2603.08814v2 Announce Type: replace Abstract: Long-horizon task planning for heterogeneous multi-robot systems is essential for deploying collaborative te…
TriPilot-FF: Coordinated Whole-Body Teleoperation with Force Feedback
arXiv:2602.09888v2 Announce Type: replace Abstract: Mobile manipulators broaden the operational envelope for robot manipulation. However, the whole-body teleope…
Cross-Modal Visuo-Tactile Representation Learning with Action Chunking Transformers for Contact-Rich Manipulation
arXiv:2602.00514v3 Announce Type: replace Abstract: Tactile feedback is important for contact-rich robotic manipulation, yet effective use of tactile observatio…
Open-Source Autonomous Driving System Analysis and Multi-Disciplinary Hardware-in-the-Loop Research Paradigm with Reinforcement-Learning Testing and Large Language Models
arXiv:2608.30179v1 Announce Type: cross Abstract: Open-source autonomous driving systems provide an inspectable software foundation for intelligent vehicle rese…
Asynchronous Cooperative Online Learning for Multi-Robot Control under Computational Delays
arXiv:2608.29562v1 Announce Type: cross Abstract: Ensuring the safe operation of multi-agent systems (MASs) under uncertain environments is crucial for cooperat…
Generalizable Multi-Agent Planning from Signal Temporal Logic Specifications via Diffusion
arXiv:2608.29490v1 Announce Type: cross Abstract: Multi-agent systems in the real-world (e.g., drone swarms, autonomous cars, warehouse robots) must satisfy ric…
CGFM-Nav: Cognitive Graph-Field Memory for Semantic-Guided Lifelong Multimodal Embodied Navigation
arXiv:2608.29114v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) requires agents to reason over accumulated observations while continuously …
Agri-Sim: Agricultural Simulation Platform for Embodied Intelligence Evaluation in Greenhouse Robotics
arXiv:2608.29100v1 Announce Type: new Abstract: Agricultural-robot development requires simulation environments that can jointly support realistic scene constru…
DREAM: Deployment-Time Demonstration Generation via Real-to-Sim for Scalable Policy Adaptation
arXiv:2608.29078v1 Announce Type: new Abstract: Vision-language-action (VLA) models have made strong progress in language-conditioned robot manipulation, but im…
LARC: Lazy Adaptive Reachability Certification of Robot Manipulator Trajectories
arXiv:2608.29767v1 Announce Type: new Abstract: Discrete trajectory checks can miss collisions between sampled robot states. Reachability-based certification bo…
SUN: Persistent Programs For Language-Grounded Control-to-Learning-to-Real Policies
arXiv:2608.31167v1 Announce Type: new Abstract: Bridging model-based control and learned policies in long-horizon manipulation has harbored a silent disagreemen…
DARP: A Calibrated Dual-Arm RGB-D-IR Dataset for Multi-View Robotic Perception
arXiv:2608.31002v1 Announce Type: new Abstract: Robotic perception from a single viewpoint is often limited by self-occlusion and incomplete surface visibility.…
A Dual-Cam Parallel Elastic Actuator with Shared Gas-Spring Compensation for Humanoid Ankles
arXiv:2608.30832v1 Announce Type: new Abstract: To improve torque capacity and energy efficiency of humanoid ankles, this paper proposes a 2-DoF parallel elasti…
Learning to infer and manipulate through distributed whole-arm interaction in a soft robot
arXiv:2608.30773v1 Announce Type: new Abstract: In animals such as elephants and octopuses, acquiring non-visual information about an object and physically enga…
Anomaly Detection on Small Industrial Components via Vision-Based Tactile Sensing
arXiv:2608.30506v1 Announce Type: new Abstract: Automated inspection of small industrial components, including sub-centimetre-scale parts where defects are geom…
PAVE: Predictive Alignment and Value-Guided Evolution for World-Action Policies
arXiv:2608.30378v1 Announce Type: new Abstract: Direct vision-language-action policies generate continuous robot actions efficiently, but standard behavior clon…
SpectraTac: A Compact Camera-Free Optical Tactile Sensor with Distributed Color Sensing
arXiv:2608.30368v1 Announce Type: new Abstract: Tactile sensing is essential for physical interaction in robotics and human--machine systems. However, combining…
Contrast-Free Autonomous Navigation of Untethered Endovascular Microrobots Using Single-Plane Fluoroscopy
arXiv:2608.30220v1 Announce Type: new Abstract: Reliable three-dimensional (3D) navigation of magnetically actuated untethered microrobots remains a major barri…
Rethinking Language's Role in Efficient VLA for Autonomous Vehicles: Toward Smarter, Trustworthy Driving
arXiv:2608.30144v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are reshaping autonomous driving (AD) by unifying perception, reasoning, and…
Training-Free Action Correction for VLA Model Failures via Language Feedback
arXiv:2608.29967v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models demonstrate strong semantic understanding yet exhibit systematic failures du…
SmoothRL: Online Reinforcement Learning During Asynchronous Execution
arXiv:2608.29768v1 Announce Type: new Abstract: Deploying robot policies in the physical world requires satisfying two fundamental desiderata: reliability and s…
DriftingVLA: Native One-Step Vision-Language-Action Generation via Per-Dimension Temporal Drifting
arXiv:2608.29749v1 Announce Type: new Abstract: Conventional flow-based vision-language-action (VLA) models support expressive continuous action generation but …
Task-Relevant Feature-Dynamics Fidelity Enables Zero-Shot Sim-to-Real Transfer for Robotic Ultrasound Scanning
arXiv:2608.29516v1 Announce Type: new Abstract: Robotic ultrasound policies operating directly on B-mode images require extensive interaction data, whereas real…
A Sliding Window Filter on the Galilean Group for Consistent Aided Inertial Navigation with Unknown Measurement Delays
arXiv:2608.29514v1 Announce Type: new Abstract: We study aided inertial navigation when the aiding sensor measurements are subject to an unknown constant delay.…
Blind Dexterity: Whole-Body Humanoid Manipulation via Pure Proprioception
arXiv:2608.29487v1 Announce Type: new Abstract: We present blind, whole-body manipulation skills on a Unitree G1 humanoid using only onboard proprioception, wit…
AdaVLA: Adaptive Step Flow Matching for Training-free Acceleration of Vision-Language-Action Models
arXiv:2608.29208v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models, built upon Vision-Language Models (VLMs), have significantly enhanced robot…
SMILE: Smooth Motion for Improved Long-Horizon VLA Execution
arXiv:2608.29432v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models reduce inference cost by executing multiple actions per call, but longer hor…
NGD-SLAM: Towards Real-Time Dynamic SLAM without GPU
arXiv:2405.07392v4 Announce Type: replace Abstract: Many existing visual SLAM methods can achieve high localization accuracy in dynamic environments by leveragi…