Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
3584 storiesPrompting Robot Teams with Natural Language
arXiv:2509.24575v2 Announce Type: replace Abstract: This paper presents a framework to prompt multi-robot teams with high-level tasks using natural language exp…
Locker-based Truck-Drone Routing with Integrated Considerations of Pickups, Deliveries, and No-Fly Zones
arXiv:2606.30680v1 Announce Type: new Abstract: Truck-drone delivery is an emerging last-mile logistics mode combining the long-haul capacity of trucks with the…
Registering the 4D Millimeter Wave Radar Point Clouds Via Generalized Method of Moments
arXiv:2508.02187v3 Announce Type: replace Abstract: 4D millimeter wave radars (4D radars) are new emerging sensors that provide point clouds of objects with bot…
Unified Structural-Hydrodynamic Modeling of Underwater Underactuated Mechanisms and Soft Robots
arXiv:2603.07939v2 Announce Type: replace Abstract: Underwater robots are widely deployed for ocean exploration and manipulation. Underactuated mechanisms are a…
Verification-Gated Agentic Mission-State Governance for Intelligent Industrial Multi-Robot Systems
arXiv:2606.31339v1 Announce Type: new Abstract: Agentic artificial intelligence is increasingly used to decompose industrial tasks, propose robot actions, and a…
Physically Grounded 3D Generative Reconstruction under Hand Occlusion using Proprioception and Multi-Contact Touch
arXiv:2604.09100v2 Announce Type: replace-cross Abstract: We propose a multimodal, physically grounded approach for metric-scale amodal object reconstruction an…
ViTL: Temporal Logic-Guided Zero-Shot Natural Language Navigation via Vision-Language Models
arXiv:2606.30696v1 Announce Type: new Abstract: Enabling robots to follow natural language commands to complete zero-shot long-horizon tasks remains challenging…
Labimus: A Simulation and Benchmark for Humanoid Dexterous Manipulation in Chemical Laboratory
arXiv:2606.31037v1 Announce Type: new Abstract: Laboratory automation has made remarkable progress through robotic platforms and AI-driven scientific reasoning.…
Long-term Traffic Simulation via Structured Autoregressive Modeling
arXiv:2606.31209v1 Announce Type: cross Abstract: Interactive traffic simulation is a vital world model for autonomous driving. A central challenge in long-hori…
VertiAdaptor: Online Kinodynamics Adaptation for Vertically Challenging Terrain
arXiv:2603.06887v3 Announce Type: replace Abstract: Autonomous driving in off-road environments presents significant challenges due to the dynamic and unpredict…
Learning Dexterous Grasping from Sparse Taxonomy Guidance
arXiv:2604.04138v2 Announce Type: replace Abstract: Dexterous manipulation requires planning a grasp configuration suited to the object and task, which is then …
Z-1: Efficient Reinforcement Learning for Vision-Language-Action Models
arXiv:2606.31846v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models offer a promising framework for robotic manipulation by connecting language …
MIRTH: Mutual-Information Reasoning with Temporal Hubs for Vision-Language-Action Agents
arXiv:2606.31167v1 Announce Type: new Abstract: VLA models have emerged as a powerful paradigm for transferring semantic knowledge from web-scale data to physic…
Sampling-Based Coordination-Informed Multi-Objective Multi-Robot Reinforcement Learning
arXiv:2606.30893v1 Announce Type: new Abstract: Multi-robot systems must simultaneously optimize competing objectives while maintaining coordinated behavior. Ex…
Autonomous UAV Navigation for Individual Wildlife Re-Identification
arXiv:2606.31772v1 Announce Type: new Abstract: Reliable individual re-identification (re-ID) of wildlife is essential for population monitoring, behavioral tra…
Towards Generalizable Robotic Manipulation in Dynamic Environments
arXiv:2603.15620v3 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models excel in static manipulation but struggle in dynamic environments …
Bridging Local Observation and Global Simulation in Closed-Loop Traffic Modeling
arXiv:2606.31844v1 Announce Type: new Abstract: A local-to-global context mismatch arises when autoregressive traffic simulators trained on ego-centric driving …
ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies
arXiv:2606.31132v1 Announce Type: new Abstract: Generative control policies (GCPs), such as diffusion policies and flow-based vision-language-action models, ena…
Diffusion-based 4D Trajectory Prediction and Distributed Control for UAV Swarms
arXiv:2606.31197v1 Announce Type: new Abstract: Accurate 4D trajectory prediction and closed-loop tracking are essential for Unmanned Aerial Vehicle (UAV) swarm…
3D HAMSTER: Bridging Planning and Control in Hierarchical Vision Language Action Models through 3D Trajectory Guidance
arXiv:2606.31329v1 Announce Type: new Abstract: Hierarchical Vision-Language-Action (VLA) models decouple high-level planning from low-level control to improve …
ChronoFlow-Policy: Unifying Past-Current-Future Interaction Flow in Visuomotor Policy Learning
arXiv:2606.31493v1 Announce Type: new Abstract: Visual signals play a crucial role in policy learning by enabling models to capture object motion and interactio…
DynFly: Dynamic-Aware Continuous Trajectory Generation for UAV Vision-Language Navigation in Urban Environments
arXiv:2606.31654v1 Announce Type: new Abstract: Recent advances in multimodal large models have significantly improved UAV vision-language navigation (UAV-VLN) …
Streaming Gaussian Encoding for 4D Panoptic Occupancy Tracking
arXiv:2606.30754v1 Announce Type: cross Abstract: Camera-based 4D panoptic occupancy tracking (4D-POT) is a promising paradigm for holistic scene understanding …
FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion
arXiv:2606.31691v1 Announce Type: new Abstract: Scalable reinforcement learning has popularized high-throughput sampling architectures, which significantly comp…
GaussLite: Online Task-Conditioned 3D Gaussian Splatting for Real-Time Robotic Mapping
arXiv:2606.30809v1 Announce Type: cross Abstract: Existing 3D Gaussian Splatting (3DGS) systems distribute representation capacity uniformly across a scene, ign…
Robust Autonomous UAV Landing on Maritime Platforms via Multimodal Agentic AI and Active Wave Compensation
arXiv:2606.31613v1 Announce Type: cross Abstract: Autonomous aerial inspection of marine infrastructure is frequently compromised by stochastic sea states, intr…
Energy-Optimal Spatial Iterative Learning within a Virtual Tube
arXiv:2606.31487v1 Announce Type: new Abstract: Due to the limited endurance of embedded energy sources such as lithium-polymer (LiPo) batteries, the flight dur…
The Quadruped Soft Tail: Compliant Grasping and Swabbing for Contamination Surveys in Harsh Environments
arXiv:2606.30900v1 Announce Type: new Abstract: Beryllium contamination surveys in radioactive areas are challenging for robots in environments cluttered with c…
Machine Learning-based Feedback Linearization Control of Quadrotor Subject to Unmodeled Dynamics
arXiv:2606.31199v1 Announce Type: new Abstract: The control of agile quadrotors in dynamic and uncertain environments remains an open area of investigation to t…
DVG-WM: Disentangled Video Generation Enables Efficient Embodied World Model for Robotic Manipulation
arXiv:2606.32028v1 Announce Type: new Abstract: Video-based embodied world models provide an appealing substrate for robotic manipulation by predicting future s…