2 citations · 2 across the 1 of their papers we have counts for
1 paper
Zhutian Yang, Caelan Garrett, Dieter Fox +2
Vision-Language Models (VLM) can generate plausible high-level plans when prompted with a goal, the context, an image of the scene, and any planning constraints. However, there is…