7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.CV2022
Understanding Attention for Vision-and-Language Tasks
Feiqi Cao, Soyeon Caren Han, Siqu Long +2
Attention mechanism has been used as an important component across Vision-and-Language(VL) tasks in order to bridge the semantic gap between visual and textual features. While atte…
cs.CV2022★ 7 cited
Vision-and-Language Pretrained Models: A Survey
Siqu Long, Feiqi Cao, Soyeon Caren Han +1
Pretrained models have produced great success in both Computer Vision (CV) and Natural Language Processing (NLP). This progress leads to learning joint representations of vision an…