1 paper · 1 filter
Junhao Shi, Zhaoye Fei, Siyin Wang +3
Large Vision-Language Models (LVLMs) show promise for embodied planning tasks but struggle with complex scenarios involving unfamiliar environments and multi-step goals. Current ap…