16 citations · 22 across the 3 of their papers we have counts for
3 papers
cs.CL2022
WikiDiverse: A Multimodal Entity Linking Dataset with Diversified Contextual Topics and Entity Types
Xuwu Wang, Junfeng Tian, Min Gui +5
Multimodal Entity Linking (MEL) which aims at linking mentions with multimodal contexts to the referent entities from a knowledge base (e.g., Wikipedia), is an essential task for m…
cs.MM2021★ 6 cited
Grid-VLP: Revisiting Grid Features for Vision-Language Pre-training
Ming Yan, Haiyang Xu, Chenliang Li +4
Existing approaches to vision-language pre-training (VLP) heavily rely on an object detector based on bounding boxes (regions), where salient objects are first detected from images…
cs.CL2019★ 16 cited
Attention Optimization for Abstractive Document Summarization
Min Gui, Junfeng Tian, Rui Wang +1
Attention plays a key role in the improvement of sequence-to-sequence-based document summarization models. To obtain a powerful attention helping with reproducing the most salient…