News
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Latest in News
4147 storiesWorldDiT: A Unified Diffusion Architecture for World and Action Modeling
arXiv:2607.23909v1 Announce Type: cross Abstract: Many recent robot policies pursue stronger control by using large pretrained vision-language models (VLMs) as …
NEO: NeRF It Once, Edit It Many Times for Continuous Object Manipulation
arXiv:2607.24538v1 Announce Type: new Abstract: In this paper, we present NEO, a unified framework providing language-guided NeRF editing for robotic manipulati…
ArmnetBench v0.1: Parallel Real-World Evaluation of Manipulation Policies on a Low-Cost Arm Farm
arXiv:2607.24481v1 Announce Type: new Abstract: Real-world evaluation is a bottleneck in developing generalist robot manipulation policies. Each rollout require…
LabRobFail: A Benchmark for Robotic Failure Analysis in Chemical Self-driving Laboratories
arXiv:2607.23704v1 Announce Type: new Abstract: The deployment of embodied agents in self-driving laboratories could accelerate scientific discovery, yet their …
Real2Sim2Real for Vision-Language-Action Manipulation: An AMD ROCm-Based Pipeline
arXiv:2607.22997v1 Announce Type: new Abstract: Physical AI -- the integration of large vision-language-action (VLA) models with embodied agents that act in the…
SimBEV2X: A Large-Scale Dataset and Data Generation Tool for Multi-Task Vehicle-to-Everything Cooperative Perception
arXiv:2607.23910v1 Announce Type: cross Abstract: Cooperative perception through vehicle-to-everything (V2X) communication can overcome the inherent physical li…
Kernel-SDF: An Open-Source Library for Real-Time Signed Distance Function Estimation using Kernel Regression
arXiv:2603.29227v2 Announce Type: replace Abstract: Accurate and efficient scene representation is crucial for robotic tasks such as motion planning, manipulati…
Low-Latency Turn-Taking via Context-Aware Preface Generation in a Real-World Dialogue Robot
arXiv:2607.23204v1 Announce Type: new Abstract: Large language model (LLM)-based dialogue systems suffer response delays because generation begins only after fi…
Data Pyramid for Embodied Manipulation
arXiv:2607.24744v1 Announce Type: new Abstract: Multimodal foundation models learned to see and to speak by consuming the whole internet. Embodied agents admit …
Learning Adaptive Multi-Task Guidance, Navigation, and Control via Hypernetworks
arXiv:2607.24292v1 Announce Type: new Abstract: Autonomous free-flying robots in orbital environments require controllers that are both versatile and resource-e…
Learning Traversability-Aware Global Planners for Long Horizon Off-Road Navigation
arXiv:2607.23743v1 Announce Type: new Abstract: Autonomous navigation across large off-road environments remains a challenging problem. Onboard sensors perceive…
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training
arXiv:2512.24125v3 Announce Type: replace Abstract: General-purpose robotic systems operating in open-world environments must achieve both broad generalization …
MAGS-SLAM: Monocular Multi-Agent Gaussian Splatting SLAM for Geometrically and Photometrically Consistent Reconstruction
arXiv:2605.10760v2 Announce Type: replace Abstract: Collaborative photorealistic 3D reconstruction from multiple agents enables rapid large-scale scene capture …
Memory for Attention: Language-Conditioned Re-Perception with a Vision--Language--Motion Map
arXiv:2607.23797v1 Announce Type: new Abstract: A robot carrying a persistent, behavior-annotated map faces two planning questions, and its memory answers only …
Cost-Aware Recovery-Pathway Identification and Bayesian Optimization for Autonomous Materials Discovery
arXiv:2607.23896v1 Announce Type: cross Abstract: Autonomous laboratories automate experimental execution, but a campaign must also decide which recovery pathwa…
Physical AI Governance: From Theory to Practice Across Life Cycle
arXiv:2607.22877v1 Announce Type: cross Abstract: With the emergence of Physical AI, artificial intelligence is extending beyond screen-based applications to em…
SHARE: Towards Head-Mounted AR with User-Centric SLAM in Shared Human-Robot Workspaces
arXiv:2607.23901v1 Announce Type: cross Abstract: Human-Robot Collaboration (HRC) in shared physical spaces using Augmented Reality (AR) interfaces is powered b…
Try Once, Then Optimal: De-Redundified Procedure Memory for Cross-Episode Exploration Amortization
arXiv:2607.23702v1 Announce Type: new Abstract: Manipulating objects with hidden internal state, such as a latched microwave, forces a robot to probe before it …
Co-planning of Flight Corridors and Communication Infrastructure for Urban Drone Logistics Networks
arXiv:2607.23989v1 Announce Type: new Abstract: Reliable wireless connectivity is essential for urban air mobility (UAM) networks in dense urban environments. I…
Learning-based Hierarchical Tracheal Anatomy Understanding from Sparse Surgical Demonstration Annotations for Ultrasound Robots
arXiv:2607.22789v1 Announce Type: cross Abstract: Tracheostomy requires precise localization of the tracheal incision site; however, conventional manual palpati…
CReF: Cross-modal and Recurrent Fusion for Depth-conditioned Humanoid Locomotion
arXiv:2603.29452v3 Announce Type: replace Abstract: Stable traversal over geometrically complex terrain increasingly requires exteroceptive perception, yet prio…
Stress-testing large language model agents in a robotic chemistry laboratory
arXiv:2607.23045v1 Announce Type: cross Abstract: AI is evaluated through knowledge, reasoning and plan generation, yet scientific agency requires reliable phys…
AutoWorld: Learning Multi-Agent Traffic Simulation with Self-Supervised World Models
arXiv:2603.28963v2 Announce Type: replace Abstract: Simulation with realistic traffic agents is essential for validating autonomous driving systems. Existing da…
DeReCo: Decoupling Representation and Coordination Learning for Object-Adaptive Decentralized Multi-Robot Cooperative Transport
arXiv:2603.08111v2 Announce Type: replace Abstract: Generalizing decentralized multi-robot cooperative transport across objects with diverse shapes and physical…
HELIOS: An LLM-Driven Autonomous Indirect Trajectory Optimization Agent
arXiv:2607.24051v1 Announce Type: cross Abstract: Low-thrust trajectory optimization is a core technology in deep-space mission design. Indirect methods based o…
A Few Words Go a Long Way: Language Guided Robot Policy Synthesis
arXiv:2607.23784v1 Announce Type: new Abstract: While vision-language-action models have demonstrated impressive zero-shot manipulation capabilities, they remai…
Moving-Horizon Estimation and Nonlinear Model Predictive Control of Cable-Driven Soft Manipulators
arXiv:2607.24029v1 Announce Type: new Abstract: Precise control of soft manipulators remains challenging due to the difficulty of developing accurate yet comput…
BC-NMPC: Battery-Constrained NMPC with Propulsion Prediction and Replanning for High-Speed Flight
arXiv:2607.23867v1 Announce Type: new Abstract: Trajectory tracking performance of Uncrewed Aerial Vehicles (UAVs) degrades during high-speed and agile flight d…
WCM: World-Cognition Model for Generalizable Human-Robot Interaction
arXiv:2607.22999v1 Announce Type: new Abstract: Language agents can now interact fluently with users in software, but robots still struggle to bring comparable …
PAC-DP: PAC-Bayesian Diffusion Policy Learning
arXiv:2607.24296v1 Announce Type: new Abstract: Diffusion Policies (DPs) are able to perform complex manipulation tasks. However, DPs are typically trained by m…