1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Hanxun Huang, Sarah Erfani, Yige Li +2
As Contrastive Language-Image Pre-training (CLIP) models are increasingly adopted for diverse downstream tasks and integrated into large vision-language models (VLMs), their suscep…