1 paper · 1 filter
Jianzhe Gao, Rui Liu, Wenguan Wang
Vision-language navigation (VLN) requires an agent to traverse complex 3D environments based on natural language instructions, necessitating a thorough scene understanding. While e…