1 paper · 1 filter
Zhenxing Xu, Brikit Lu, Yihong Lu +10
Current Visual-Language Navigation (VLN) methodologies face a trade-off between semantic understanding and control precision. While Multimodal Large Language Models (MLLMs) offer s…