1 paper
Zhenxing Xu, Brikit Lu, Yihong Lu +10
Current Visual-Language Navigation (VLN) methodologies face a trade-off between semantic understanding and control precision. While Multimodal Large Language Models (MLLMs) offer s…