Xiaomi-Robotics-1 overcomes physical robotics data bottlenecks by utilizing 100,000 hours of embodiment-free pre-training data aligned via a two-stage training pipeline. The resulting vision-language-action model demonstrates predictable performance scaling and state-of-the-art success rates in both physical deployment and simulation environments.