1 paper
Xiaoxing Hu, Kaicheng Yang, Jun Wang +3
Contrastive Language-Image Pre-training (CLIP) has achieved success on multiple downstream tasks by aligning image and text modalities. However, the nature of global contrastive le…