1 paper · 1 filter
Hanyao Wang, Yibing Zhan, Liu Liu +3
Pretrained cross-modal models, for instance, the most representative CLIP, have recently led to a boom in using pre-trained models for cross-modal zero-shot tasks, considering the…