1 paper
Junhao Dong, Yifei Zhang, Hao Zhu +2
Vision-Language Models (VLMs) can perform zero-shot classification but are susceptible to adversarial attacks. While robust fine-tuning improves their robustness, existing approach…