12 citations · 17 across the 6 of their papers we have counts for
1 paper · 1 filter
Hyunjae Kim, Seunghyun Yoon, Trung Bui +4
Contrastive language-image pre-training (CLIP) models have demonstrated considerable success across various vision-language tasks, such as text-to-image retrieval, where the model…