News
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Daily robotics industry news: humanoids, industrial automation, drones, autonomous systems and the companies building them.
Latest in News
2456 storiesTDMA Based Communications Control Co-Design for Cooperative Carrying: Delay Calibration and Sampling-Rate Optimization
arXiv:2608.09556v1 Announce Type: new Abstract: Multi robot teams performing cooperative transportation face a fundamental challenge: maintaining stable control…
FactorDrive: Adaptive Multi-Step Reasoning Driven by Planning-Critical Factors for End-to-End Autonomous Driving
arXiv:2608.09591v1 Announce Type: new Abstract: Vision-language models (VLMs) have advanced scene understanding and enabled explicit reasoning in end-to-end aut…
Nonlinear Model Predictive Control of a Robotic Soft Esophagus
arXiv:2608.09602v1 Announce Type: new Abstract: Strictures caused by esophageal cancer can narrow down the esophageal lumen, leading to dysphagia. Palliation of…
TAMS: Task-Aware Multi-View Adaptive Streaming for Wireless Telerobotic Manipulation
arXiv:2608.09731v1 Announce Type: new Abstract: Wireless telerobotic manipulation relies on timely multi-view video feedback, but the available uplink bandwidth…
Efficient Real-World Online Reinforcement Learning for Robot Manipulation via Centralized Training and Critic Decomposition
arXiv:2608.09762v1 Announce Type: new Abstract: Real-world online reinforcement learning (RL) provides a promising approach for training robotic manipulation po…
SLIM-0.5B: Learning Action-Grounded Predictive Latents for Robot Manipulation
arXiv:2608.09771v1 Announce Type: new Abstract: Vision-language-action policies rely on large multimodal backbones to jointly perform perception, language condi…
RoboSeg: Online Part-Level Semantic Reconstruction for Robotic Manipulation via a Single Eye-in-Hand Camera
arXiv:2608.09778v1 Announce Type: new Abstract: Robotic manipulation requires perception systemsthat identify actionable parts such as handles, rims, triggers,a…
WRAP: Wasserstein-Robust Adaptive Plug-in for Robot Localization
arXiv:2608.09807v1 Announce Type: new Abstract: Robotic localization under changing sensing conditions can suffer from biased errors and miscalibrated covarianc…
Hierarchical Fast--Slow ReAct Agent for Zero-Shot Object-Goal Navigation
arXiv:2608.09816v1 Announce Type: new Abstract: Zero-shot object-goal navigation (ZSON) requires a robot to find a named object category in a building it has ne…
Agentic Harnesses: LLM-Driven Verification Layers for Robot Autonomy
arXiv:2608.09857v1 Announce Type: new Abstract: Advances in advanced artificial intelligence tools have sparked research in robot autonomy, but the development …
Energy-Structured Latent World Models with Neural Time Fields for Physically Constistent Open-World Motion Planning
arXiv:2608.09876v1 Announce Type: new Abstract: Physically consistent motion planning remains a fundamental challenge in embodied AI, as generated trajectories …
RoSE: A Robotic Soft Esophagus for Endoprosthetic Stent Testing
arXiv:2608.09891v1 Announce Type: new Abstract: Soft robotic systems are well suited for developing devices for biomedical applications. A bio-mimicking robotic…
XPolicyLab: A Unified Standard and Open Ecosystem for Robot Policy Evaluation and Deployment
arXiv:2608.09892v1 Announce Type: new Abstract: Robot policy evaluation and deployment remain fragmented by model-specific software dependencies, data represent…
An AI Scientist that Doesn't Drift: Taste, Structure, and Falsifiable Findings in a Quadruped Navigation Research Loop
arXiv:2608.07542v1 Announce Type: cross Abstract: Autonomous research loops driven by large language models can run machine-learning experiments at scale but te…
The Field Knows: Cross-Dimensional Geometry from Navigation to Black Holes
arXiv:2608.07566v1 Announce Type: cross Abstract: We introduce a continuous metric field framework trained by a single causal contrastive loss. The framework en…
Open-World Hierarchical Perception: Taxonomic Abstraction over Class-Agnostic Proposals for the Safe Handling of Out-of-Vocabulary Road Objects
arXiv:2608.07577v1 Announce Type: cross Abstract: A closed-set detector for autonomous driving must assign every object one of a fixed set of labels. On an obje…
CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark and Vehicle-to-Vehicle Vision-Language-Action Models
arXiv:2608.07621v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have recently achieved impressive performance for end-to-end autonomous dr…
LUCID: Latent-Skill Unified Control via Imagined Dynamics for Long-Horizon Humanoid Loco-Manipulation
arXiv:2608.07746v1 Announce Type: cross Abstract: Long-horizon humanoid loco-manipulation requires composing versatile whole-body skills and reliable high-level…
Explore, Map, Remember, Decide: Are Embodied VLMs Ready for Safety-Critical Scenarios?
arXiv:2608.08077v1 Announce Type: cross Abstract: Theory of Space framework (ToS) assesses the spatial understanding of curiosity-driven Vision-Language Models …
Preview-Based Relative-Motion Control of an Insertion Tool for Neural-Thread Placement in Pulsating Tissue
arXiv:2608.08860v1 Announce Type: cross Abstract: Robotic neural-thread placement requires regulating the insertion-tool tip relative to tissue that moves with …
Real-Time Nonlinear MPC via Sequential Quadratic Programming with Structure-Exploiting ADMM and Interior-Point Methods for Underactuated Double-Pendulum Swing-Up
arXiv:2608.09272v1 Announce Type: cross Abstract: The 4th "AI Olympics with RealAIGym" competition, to be held at IJCAI-ECAI 2026 in Bremen, challenges particip…
A Height-Constrained 2-Point Minimal Solver for Pose Estimation from Active LED Markers with Event Cameras
arXiv:2608.09520v1 Announce Type: cross Abstract: In many autonomous applications requiring real-time localization, active marker-based systems are preferred du…
A Semantic Communication Approach to Fiducial Marker Processing in 5G-Enabled Edge SLAM
arXiv:2608.09620v1 Announce Type: cross Abstract: Autonomous robots increasingly rely on edge computing to offload computationally intensive perception tasks wh…
REMAC: Self-Reflective and Self-Evolving Multi-Agent Collaboration for Long-Horizon Robot Manipulation
arXiv:2503.22122v2 Announce Type: replace Abstract: Vision-language models (VLMs) have demonstrated remarkable capabilities in robotic planning, particularly fo…
X2C: A Dataset Featuring Nuanced Facial Expressions for Realistic Humanoid Imitation
arXiv:2505.11146v3 Announce Type: replace Abstract: Fine-grained facial expression transfer from humans to humanoid agents presents a unique pattern recognition…
Calib3R: Hand-Eye Calibration and 3D Metric-Scaled Scene Reconstruction with 3D Foundation Models
arXiv:2509.08813v2 Announce Type: replace Abstract: Robots often rely on RGB images for tasks like manipulation. However, reliable interaction typically require…
CLEAR: A Semantic-Geometric Terrain Abstraction for Large-Scale Unstructured Environments
arXiv:2601.13361v2 Announce Type: replace Abstract: Long-horizon navigation in unstructured environments demands terrain abstractions that scale to tens of squa…
HyperDet: 3D Object Detection with Hyper 4D Radar Point Clouds
arXiv:2602.11554v4 Announce Type: replace Abstract: How far can 3D object detection go using 4D radar alone? Despite offering weather-robust and velocity-aware …
AnyCamVLA: Zero-Shot Camera Adaptation for Viewpoint Robust Vision-Language-Action Models
arXiv:2603.05868v2 Announce Type: replace Abstract: Despite remarkable progress in Vision-Language-Action models (VLAs) for robot manipulation, these large pre-…
Knowing When to Ask: Resolving Uncertainty in Human-Robot Joint Planning via Explicit Dialogue and Implicit Intent Cues
arXiv:2603.07822v2 Announce Type: replace Abstract: Effective human-robot collaboration in open-world environments requires joint planning under uncertainty abo…