3 citations · 3 across the 11 of their papers we have counts for
1 paper · 2 filters
Qi Qian, Yuanhong Xu, Juhua Hu
Vision-language pre-training methods, e.g., CLIP, demonstrate an impressive zero-shot performance on visual categorizations with the class proxy from the text embedding of the clas…