54 citations · 54 across the 1 of their papers we have counts for
1 paper
Tony Huang, Jack Chu, Fangyun Wei
Contrastive vision-language models like CLIP have shown great progress in transfer learning. In the inference stage, the proper text description, also known as prompt, needs to be…