Research
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Frontier robotics research: arXiv cs.RO papers, embodied AI, VLA models, manipulation, navigation and learning systems.
Latest in Research
3542 storiesAdaptive-Horizon Conflict-Based Search for Closed-Loop Multi-Agent Path Finding
arXiv:2602.12024v2 Announce Type: replace Abstract: MAPF is a core coordination problem for large robot fleets in automated warehouses and logistics. Existing a…
AssemLM: A Spatial Reasoning Multimodal Large Language Model for Robotic Assembly
arXiv:2604.08983v2 Announce Type: replace Abstract: Spatial reasoning is a fundamental capability for embodied intelligence, especially for fine-grained manipul…
From Digital to Physical: Digital Agents as Autonomous Coaches for Physical Intelligence
arXiv:2601.21570v2 Announce Type: replace-cross Abstract: The field of Embodied AI is witnessing a rapid evolution toward general-purpose robotic systems, fuele…
Triangle Splatting SLAM
arXiv:2605.31419v2 Announce Type: replace-cross Abstract: We present a dense RGB-D SLAM system using differentiable triangles as the 3D map representation. Whil…
Goal2Pixel: Grounding Goals to Pixels for Vision-Language Navigation
arXiv:2606.01621v2 Announce Type: replace-cross Abstract: Vision-language models (VLMs) have become a common foundation for vision-and-language navigation in co…
From Seeing to Experiencing: Scaling Navigation Foundation Models with Reinforcement Learning
arXiv:2507.22028v2 Announce Type: replace-cross Abstract: Navigation foundation models trained on massive web-scale data enable agents to generalize across dive…
Action-Effect Memory Pretraining for Robot Manipulation
arXiv:2606.12499v1 Announce Type: new Abstract: We present AEM, an Action-Effect Memory pretraining framework for robot manipulation that learns compact tempora…
EgoMoD: Predicting Global Maps of Dynamics from Local Egocentric Observations
arXiv:2603.00167v2 Announce Type: replace Abstract: Efficient navigation in dynamic environments requires anticipating how motion patterns evolve beyond the rob…
Miniature Testbed for Validating Multi-Agent Cooperative Autonomous Driving
arXiv:2511.11022v2 Announce Type: replace Abstract: Cooperative autonomous driving, which extends vehicle autonomy by enabling real-time collaboration between v…
GLIDE: A Coordinated Aerial-Ground Framework for Search and Rescue in Unknown Environments
arXiv:2509.14210v4 Announce Type: replace Abstract: We present a cooperative aerial-ground search-and-rescue (SAR) framework that pairs two unmanned aerial vehi…
Data-Driven Soft Robot Control via Adiabatic Spectral Submanifolds
arXiv:2503.10919v3 Announce Type: replace Abstract: The mechanical complexity of soft robots creates significant challenges for their model-based control. Speci…
PolyFlow: Safe and Efficient Polytope-Constrained Flow Matching with Constraint Embedding and Projection-free Update
arXiv:2606.13400v1 Announce Type: cross Abstract: While flow-based generative models have demonstrated strong performance across a wide range of domains, deploy…
Visual Place Recognition in Forests with Depth-Aware Distillation
arXiv:2606.13206v1 Announce Type: cross Abstract: Visual place recognition in natural forest environments remains challenging due to repetitive vegetation, weak…
MPC for underactuated spacecraft control with a Lyapunov supervised physics-informed neural network correction layer
arXiv:2606.13113v1 Announce Type: cross Abstract: Underactuated spacecraft faces controllability limitations and heightened sensitivity to environmental disturb…
Effects of Social Interactions in Self-Organising Railway Traffic Management
arXiv:2606.13068v1 Announce Type: cross Abstract: Recent research is exploring self-organised traffic management as a solution for scaling to complex real-world…
TrajGenAgent: A Hierarchical LLM Agent for Human Mobility Trajectory Generation
arXiv:2606.12657v1 Announce Type: cross Abstract: Human mobility data is important for transportation, urban planning, and epidemic control, but large-scale tra…
Individual Control Barrier Functions-Guided Diffusion Model for Safe Offline Multi-Agent Reinforcement Learning
arXiv:2606.12640v1 Announce Type: cross Abstract: Offline reinforcement learning allows control policies to be learned directly from data without online interac…
WOMBET: World Model-Based Experience Transfer for Robust and Sample-efficient Reinforcement Learning
arXiv:2604.08958v3 Announce Type: replace-cross Abstract: Reinforcement learning (RL) in robotics is often limited by the cost and risk of data collection, moti…
GAE: Unleashing Physical Potential of VLM with Generalizable Action Expert
arXiv:2510.03896v2 Announce Type: replace-cross Abstract: Vision-language models demonstrate strong reasoning and planning abilities, yet grounding these predic…
DiffCoord: Differentiable Coordination for Distributed Multi-Agent Trajectory Optimization
arXiv:2509.01630v3 Announce Type: replace-cross Abstract: Integrating the Alternating Direction Method of Multipliers (ADMM) with Differential Dynamic Programmi…
Safety Case Patterns for VLA-based driving systems: Insights from SimLingo
arXiv:2603.16013v3 Announce Type: replace Abstract: Vision-Language-Action (VLA)-based driving systems represent a significant paradigm shift in autonomous driv…
PROBE: Probabilistic Occupancy BEV Encoding with Analytical Translation Robustness for 3D Place Recognition
arXiv:2603.05965v3 Announce Type: replace Abstract: We present PROBE (PRobabilistic Occupancy BEV Encoding), a learning-free LiDAR place recognition descriptor …
Lyapunov-Based PI-Like Control for Robust Trajectory Tracking of a Four-Wheel Independently Driven and Steered Robot: Design and Experimental Validation
arXiv:2602.15424v2 Announce Type: replace Abstract: In this paper, a Lyapunov-based synthesis of a PI-like controller is proposed for robust trajectory tracking…
Extending the Law of Intersegmental Coordination: Implications for Powered Prosthetic Controls
arXiv:2602.02181v2 Announce Type: replace Abstract: Powered prostheses are capable of providing net positive work to amputees and have advanced in the past two …
DiskChunGS: Large-Scale 3D Gaussian SLAM Through Chunk-Based Memory Management
arXiv:2511.23030v2 Announce Type: replace Abstract: Recent advances in 3D Gaussian Splatting (3DGS) have demonstrated impressive results for novel view synthesi…
Learning Visually Interpretable Oscillator Networks for Soft Continuum Robots from Video
arXiv:2511.18322v4 Announce Type: replace Abstract: Learning soft continuum robot (SCR) dynamics from video offers flexibility but existing methods lack interpr…
Adaptive Model-Predictive Control of a Soft Continuum Robot Using a Physics-Informed Neural Network Based on Cosserat Rod Theory
arXiv:2508.12681v3 Announce Type: replace Abstract: Dynamic control of soft continuum robots (SCRs) holds great potential for expanding their applications, but …
Heterogeneous LiDAR Early Fusion and Learned Re-Ranking Strategy for Robust Long-Term Place Recognition in Unstructured Environments
arXiv:2606.13503v1 Announce Type: cross Abstract: Robust localization in unstructured environments, such as agricultural fields, is a critical challenge for aut…
Diffusion Transformer World-Action Model for AV Scene Prediction
arXiv:2606.12987v1 Announce Type: cross Abstract: Action-conditioned world models let an autonomous vehicle predict future camera scenes from its own planned co…
SemanticXR: Low Power and Real-time Queryable Semantic Mapping with an Object-Level Device-Cloud Architecture
arXiv:2606.12849v1 Announce Type: cross Abstract: Semantic mapping is a core service that enables grounded interactions in emerging Extended Reality (XR) applic…