1 paper
Zhenyu Zhang, Guangyao Chen, Yixiong Zou +2
The Contrastive Language-Image Pre-Training (CLIP) model excels in few-shot learning by aligning visual and textual representations. Our study shows that template-sample similarity…