1 citations · 1 across the 1 of their papers we have counts for
1 paper
Wenqi Lyu, Zerui Li, Yanyuan Qiao +1
Multimodal large language models (MLLMs) have recently gained attention for their generalization and reasoning capabilities in Vision-and-Language Navigation (VLN) tasks, leading t…