7 citations · 7 across the 1 of their papers we have counts for
1 paper
Xinyu Huang, Youcai Zhang, Ying Cheng +7
Vision-Language Pre-training (VLP) with large-scale image-text pairs has demonstrated superior performance in various fields. However, the image-text pairs co-occurrent on the Inte…