2 citations · 2 across the 1 of their papers we have counts for
1 paper · 1 filter
Jie-Jing Shao, Jiang-Xin Shi, Xiao-Wen Yang +2
Contrastive Language-Image Pre-training (CLIP) provides a foundation model by integrating natural language into visual concepts, enabling zero-shot recognition on downstream tasks.…