1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Chenyu You, Yifei Min, Weicheng Dai +3
Fine-tuning pre-trained vision-language models, like CLIP, has yielded success on diverse downstream tasks. However, several pain points persist for this paradigm: (i) directly tun…