94 citations · 226 across the 4 of their papers we have counts for
1 paper · 2 filters
Zihao Zhao, Yuxiao Liu, Han Wu +8
Contrastive Language-Image Pre-training (CLIP), a simple yet effective pre-training paradigm, successfully introduces text supervision to vision models. It has shown promising resu…