1 paper · 1 filter
Ruijie Zheng, Yongyuan Liang, Shuaiyi Huang +5
Although large vision-language-action (VLA) models pretrained on extensive robot datasets offer promising generalist policies for robotic learning, they still struggle with spatial…