13 citations · 13 across the 1 of their papers we have counts for
2 papers
cs.IR2021★ 13 cited
MURAL: Multimodal, Multitask Retrieval Across Languages
Aashi Jain, Mandy Guo, Krishna Srinivasan +5
Both image-caption pairs and translation pairs provide the means to learn deep representations of and connections between languages. We use both types of pairs in MURAL (MUltimodal…
cs.CV2021
Scaling Up Visual and Vision-Language Representation Learning With Noisy Text Supervision
Chao Jia, Yinfei Yang, Ye Xia +7
Pre-trained representations are becoming crucial for many NLP and perception tasks. While representation learning in NLP has transitioned to training on raw text without human anno…