1 citations · 2 across the 3 of their papers we have counts for
1 paper · 1 filter
Hao Tan, Jun Li, Yizhuang Zhou +3
Vision-Language Models (VLMs) such as CLIP have demonstrated remarkable generalization capabilities to downstream tasks. However, existing prompt tuning based frameworks need to pa…