1 paper · 1 filter
Xiaohan Zhang, Zainab Altaweel, Yohei Hayamizu +6
Vision-language models (VLMs) have been applied to robot task planning problems, where the robot receives a task in natural language and generates plans based on visual inputs. Whi…