1 paper
Xiangyan Qu, Gaopeng Gou, Jiamin Zhuang +5
Vision-language models (VLMs) have made significant progress in image classification by training with large-scale paired image-text data. Their performances largely depend on the p…