1 paper · 1 filter
Yihao Wu, He Zhang, Junbo Tan +2
Post-training Vision-Language-Action (VLA) models into policies that can be reliably deployed on real robots remains a major bottleneck. SFT and DAgger exploit failure signals only…