1 paper
Zihao Zhao, Yuxiao Liu, Han Wu +8
Contrastive Language-Image Pre-training (CLIP), a simple yet effective pre-training paradigm, successfully introduces text supervision to vision models. It has shown promising resu…