3 citations · 3 across the 1 of their papers we have counts for
1 paper
Bingqian Lin, Yi Zhu, Zicong Chen +3
Vision-Language Navigation (VLN) is a challenging task that requires an embodied agent to perform action-level modality alignment, i.e., make instruction-asked actions sequentially…