6 citations · 13 across the 6 of their papers we have counts for
1 paper · 1 filter
Hao Yang, Can Gao, Hao Líu +3
Vision-and-language (VL) pre-training, which aims to learn a general representation of image-text pairs that can be transferred to various vision-and-language tasks. Compared with…