4 citations · 4 across the 1 of their papers we have counts for
1 paper
Quan Sun, Jinsheng Wang, Qiying Yu +4
Scaling up contrastive language-image pretraining (CLIP) is critical for empowering both vision and multimodal models. We present EVA-CLIP-18B, the largest and most powerful open-s…