21 citations · 48 across the 9 of their papers we have counts for
1 paper · 1 filter
Jiaxing Liu, Zexi Zhang, Xiaoyan Li +3
Vision-Language Navigation (VLN) presents a unique challenge for Large Vision-Language Models (VLMs) due to their inherent architectural mismatch: VLMs are primarily pretrained on…