2 papers
cs.CV2023
DAP: Domain-aware Prompt Learning for Vision-and-Language Navigation
Ting Liu, Yue Hu, Wansen Wu +3
Following language instructions to navigate in unseen environments is a challenging task for autonomous embodied agents. With strong representation capabilities, pretrained vision-…
cs.CV2023
Prompt-based Context- and Domain-aware Pretraining for Vision and Language Navigation
Ting Liu, Yue Hu, Wansen Wu +3
Pretrained visual-language models have extensive world knowledge and are widely used in visual and language navigation (VLN). However, they are not sensitive to indoor scenarios fo…