1 paper
Xiaozhen Qiao, Jingkai Zhao, Yuqiu Jiang +4
Vision-Language Models (VLMs) demonstrate impressive zero-shot generalization through large-scale image-text pretraining, yet their performance can drop once the deployment distrib…