1 paper
Zihua Wang, Zhitao Lin, Ruibo Li +4
Vision-Language-Action (VLA) models, as large foundation models for embodied control, have shown strong performance in manipulation tasks. However, their performance comes at high…