1 paper
Nguyen Duc Thai, Junhao Dong, Sua Qi Rong +2
Vision-language models (VLMs), such as CLIP, are vulnerable to adversarial attacks, posing a serious problem for real-life applications and deployment. Adversarial fine-tuning emer…