1 paper · 1 filter
Xianghui Wang, Xinming Zhang, Yanjun Chen +2
Vision-language models (VLMs) have demonstrated excellent high-level planning capabilities, enabling locomotion skill learning from video demonstrations without the need for meticu…