3 citations · 3 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 3 cited
CPL: Counterfactual Prompt Learning for Vision and Language Models
Xuehai He, Diji Yang, Weixi Feng +7
Prompt tuning is a new few-shot transfer learning technique that only tunes the learnable prompt for pre-trained vision and language models such as CLIP. However, existing prompt t…
cs.CV2022
ULN: Towards Underspecified Vision-and-Language Navigation
Weixi Feng, Tsu-Jui Fu, Yujie Lu +1
Vision-and-Language Navigation (VLN) is a task to guide an embodied agent moving to a target position using language instructions. Despite the significant performance improvement,…
cs.CV2022
Anticipating the Unseen Discrepancy for Vision and Language Navigation
Yujie Lu, Huiliang Zhang, Ping Nie +4
Vision-Language Navigation requires the agent to follow natural language instructions to reach a specific target. The large discrepancy between seen and unseen environments makes i…