1 paper · 1 filter
Siqi Wang, Chao Liang, Yunfan Gao +5
Vision-Language Models (VLMs) have made significant progress in explicit instruction-based navigation; however, their ability to interpret implicit human needs (e.g., "I am thirsty…