code retrieval 1coding agents 1context allocation 1latent space interventions 1model robustness 1redundancy reduction 1representation analysis 1retrieval-augmented generation 1vision-language models 1visual grounding 1
From the 2 of 15 linked papers with an AI index.
Showing cs.ROShow all
2 papers · 1 filter
cs.RO2026
Xiaomi-Robotics-1: Scaling Vision-Language-Action Models with over 100K Hours of Real-World Trajectories
Xiaomi Robotics Team, Jun Guo, Piaopiao Jin +31
We present Xiaomi-Robotics-1, a foundational vision-language-action (VLA) model capable of (1) following diverse language instructions to perform a wide range of mobile manipulatio…
cs.RO2026
Xiaomi-Robotics-0: An Open-Sourced Vision-Language-Action Model with Real-Time Execution
Rui Cai, Jun Guo, Xinze He +20
In this report, we introduce Xiaomi-Robotics-0, an advanced vision-language-action (VLA) model optimized for high performance and fast and smooth real-time execution. The key to ou…