19 citations · 22 across the 2 of their papers we have counts for
4 papers
ADAPT: Vision-Language Navigation with Modality-Aligned Action Prompts
Bingqian Lin, Yi Zhu, Zicong Chen +3
Vision-Language Navigation (VLN) is a challenging task that requires an embodied agent to perform action-level modality alignment, i.e., make instruction-asked actions sequentially…
Adversarial Reinforced Instruction Attacker for Robust Vision-Language Navigation
Bingqian Lin, Yi Zhu, Yanxin Long +3
Language instruction plays an essential role in the natural language grounded navigation tasks. However, navigators trained with limited human-annotated instructions may have diffi…
Vision-Dialog Navigation by Exploring Cross-modal Memory
Yi Zhu, Fengda Zhu, Zhaohuan Zhan +4
Vision-dialog navigation posed as a new holy-grail task in vision-language disciplinary targets at learning an agent endowed with the capability of constant conversation for help w…
Jointly Deep Multi-View Learning for Clustering Analysis
Bingqian Lin, Yuan Xie, Yanyun Qu +2
In this paper, we propose a novel Joint framework for Deep Multi-view Clustering (DMJC), where multiple deep embedded features, multi-view fusion mechanism and clustering assignmen…