2 papers
cs.RO2026
SOP: A Scalable Online Post-Training System for Vision-Language-Action Models
Mingjie Pan, Siyuan Feng, Qinglin Zhang +9
Vision-language-action (VLA) models achieve strong generalization through large-scale pre-training, but real-world deployment requires expert-level task proficiency in addition to…
cs.RO2026
Unified Embodied VLM Reasoning with Robotic Action via Autoregressive Discretized Pre-training
Yi Liu, Sukai Wang, Dafeng Wei +10
General-purpose robotic systems operating in open-world environments must achieve both broad generalization and high-precision action execution, a combination that remains challeng…