6 citations · 7 across the 3 of their papers we have counts for
3 papers
cs.CV2022★ 6 cited
HiVLP: Hierarchical Vision-Language Pre-Training for Fast Image-Text Retrieval
Feilong Chen, Xiuyi Chen, Jiaxin Shi +3
In the past few years, the emergence of vision-language pre-training (VLP) has brought cross-modal retrieval to a new era. However, due to the latency and computation demand, it is…
cs.CV2022
Improving Cross-Modal Understanding in Visual Dialog via Contrastive Learning
Feilong Chen, Xiuyi Chen, Shuang Xu +1
Visual Dialog is a challenging vision-language task since the visual dialog agent needs to answer a series of questions after reasoning over both the image content and dialog histo…
cs.CL2021★ 1 cited
Multimodal Incremental Transformer with Visual Grounding for Visual Dialogue Generation
Feilong Chen, Fandong Meng, Xiuyi Chen +2
Visual dialogue is a challenging task since it needs to answer a series of coherent questions on the basis of understanding the visual environment. Previous studies focus on the im…