94 citations · 173 across the 3 of their papers we have counts for
1 paper · 1 filter
Zihao Zhao, Yuxiao Liu, Han Wu +8
Contrastive Language-Image Pre-training (CLIP), a simple yet effective pre-training paradigm, successfully introduces text supervision to vision models. It has shown promising resu…