4 papers · 1 filter
All-day Multi-scenes Lifelong Vision-and-Language Navigation with Tucker Adaptation
Xudong Wang, Gan Li, Zhiyu Liu +3
Deploying vision-and-language navigation (VLN) agents requires adaptation across diverse scenes and environments, but fine-tuning on a specific scenario often causes catastrophic f…
Exploring the Use of VLMs for Navigation Assistance for People with Blindness and Low Vision
Yu Li, Yuchen Zheng, Giles Hamilton-Fletcher +6
This paper investigates the potential of vision-language models (VLMs) to assist people with blindness and low vision (pBLV) in navigation tasks. We evaluate state-of-the-art close…
Distillation Improves Visual Place Recognition for Low Quality Images
Anbang Yang, Ge Jin, Junjie Huang +3
Real-time visual localization often utilizes online computing, for which query images or videos are transmitted to remote servers for visual place recognition (VPR). However, limit…
A Multi-Modal Foundation Model to Assist People with Blindness and Low Vision in Environmental Interaction
Yu Hao, Fan Yang, Hao Huang +5
People with blindness and low vision (pBLV) encounter substantial challenges when it comes to comprehensive scene recognition and precise object identification in unfamiliar enviro…