1 paper
Li Xian, Mingxi Li, Yizheng Wang +3
Vision-Language Navigation (VLN) requires an embodied agent to interpret a natural-language instruction and predict actions from temporally ordered visual observations. Adapting a…