1 paper
Zhihao Zhu, Yifan Zheng, Siyu Pan +2
The fragmentation between high-level task semantics and low-level geometric features remains a persistent challenge in robotic manipulation. While vision-language models (VLMs) hav…