1 citations · 1 across the 1 of their papers we have counts for
1 paper
Ali Abdollah, Amirmohammad Izadi, Armin Saghafian +5
Vision-language models (VLMs) like CLIP have showcased a remarkable ability to extract transferable features for downstream tasks. Nonetheless, the training process of these models…