1 citations · 1 across the 1 of their papers we have counts for
1 paper
Shentong Mo, Jingfei Xia, Ihor Markevych
Visual and linguistic pre-training aims to learn vision and language representations together, which can be transferred to visual-linguistic downstream tasks. However, there exists…