3 citations · 13 across the 10 of their papers we have counts for
1 paper · 1 filter
Siddharth Joshi, Arnav Jain, Ali Payani +1
Contrastive Language-Image Pre-training (CLIP) on large-scale image-caption datasets learns representations that can achieve remarkable zero-shot generalization. However, such mode…