Robotics

cs.RO

Autonomous systems, robot learning, motion planning, and manipulation.

Sort:

XPolicyLab Cuts Robot Policy Integration From NM to N Plus M

A shared policy contract and isolated client/server runtime let 42 robot policies connect to simulations and real-robot evaluation without pairwise adapters.

Aug 18, 20264 min2608.09892

Monocular Navigation Surpasses Depth-Based Systems on R2R-CE

An 8B vision-language model predicts image-space waypoints from a single RGB stream, reaching 77.4% unseen-environment success while cutting supervised training tokens 22×.

Aug 5, 20264 min2607.20785

Open Egocentric Data Lowers Barriers for Robot Learning

Smartphone capture and a modular processing toolchain turn 2,000 hours of human manipulation video into structured supervision for embodied models.

Jul 29, 20265 min2607.14183

Xiaomi Scales Robot Policies With Real Trajectories

A two-stage VLA recipe combines 100K hours of UMI trajectories with cross-embodiment post-training, reaching 57.4% on RoboCasa365.

Jul 28, 20265 min2607.15330

ACME Broadens Social Navigation Data Beyond Single-Site Robots

Eight sites across five countries and seven robot embodiments add 72.1K verified trajectories for testing navigation and prediction under varied crowds.

Jul 28, 20265 min2607.21964

Slow-Fast Navigation Model Improves Urban POI Arrival

Explicit reasoning and pixel-goal anchors decouple cognition from control, raising POI arrival to 77.3% with a reported 35.0% gain.

Jul 28, 20264 min2607.10383

OmniDreams Runs Generative AV Simulation in Real Time

A Cosmos-derived causal diffusion model uses simulator state, action cues, and a streaming KV cache to render 704×1280 rollouts at 68–105 FPS.

Jul 28, 20265 min2606.03159 Code available

ACE-Brain Extends Spatial Models Into Closed-Loop Robot Agents

An 8B backbone couples perception, planning, action, and progress estimation, improving 14 of 18 spatial benchmarks over ACE-Brain-0.

Jul 12, 20265 min2607.04426

Foundation Models for Robotics Show Strengths and Gaps

A benchmark of VLA models across 18 tasks reveals strong object generalization but weak force control and long-horizon planning.

Apr 10, 20249 min2404.05678 Code available