1 paper
Zhixuan Shen, Jiawei Du, Ziyu Guo +5
Vision-Language Models (VLMs) have demonstrated exceptional general reasoning capabilities. However, their performance in embodied navigation remains hindered by a scarcity of alig…