1 paper
Yuqi Lin, Minghao Chen, Kaipeng Zhang +7
Contrastive Language-Image Pre-training (CLIP) has demonstrated impressive capabilities in open-vocabulary classification. The class token in the image encoder is trained to captur…