Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories
We present Xiaomi-Robotics-1, a foundational vision-language-action (VLA) model capable of (1) following diverse language instructions to perform a wide range of mobile manipulation tasks in unseen environments out-of-the-box, and (2) efficiently adapting to novel downstream tasks with minimal fine-tuning data. We propose a two-stage training recipe consisting of pre-training and post-training. During pre-training, we imbue the model with broad and generalizable action-generation capabilities by
Record details
Published: 15 July 2026
Source: HuggingFace Daily Papers
Category: Research
Topics: Agents & autonomy · Environment
Retrieved: 20 July 2026
Related evidence
These records share source-supplied organisations, an exact publisher byline, automatic topics or regions. The reason is shown on every link; related does not mean supporting, agreeing with or verifying this record.
Multi-Turn On-Policy Distillation with Prefix Replay
HuggingFace Daily Papers · 15 July 2026
SEED: Self-Evolving On-Policy Distillation for Agentic Reinforcement Learning
HuggingFace Daily Papers · 15 July 2026
AI-Native Insurance for Agentic AI: Pricing, Underwriting, and End-to-End Automation
arXiv cs.CY · 16 July 2026
From Language to Navigation Goals: A Vision-Language Approach for Semantic Navigation of Mobile Robots Using RGB-D Perception
arXiv cs.AI · 15 July 2026
UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following
arXiv cs.AI · 15 July 2026
SAFETY SENTRY: Context-Aware Human Intervention via EXECUTE-ASK-REFUSE Routing
arXiv cs.AI · 15 July 2026
How to cite this record
ethics.ai (15 July 2026), “Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories,” evidence record 11743, https://ethics.ai/record/11743 (originally published by HuggingFace Daily Papers).
Use and limitations
This page is a stable index and citation surface for a source record. ethics.ai did not author the underlying report and has not independently verified every claim. Automatic topics may be imperfect. For consequential use, quote and cite the original publisher.