1 paper · 1 filter
Qingxuan Jia, Guoqin Tang, Zeyuan Huang +5
Vision-Language Models (VLMs) demonstrate remarkable potential in robotic manipulation, yet challenges persist in executing complex fine manipulation tasks with high speed and prec…