24 citations · 24 across the 2 of their papers we have counts for
2 papers
cs.CV2026
Adversarial Attacks Already Tell the Answer: Directional Bias-Guided Test-time Defense for Vision-Language Models
Liangsheng Liu, Si Chen, Jiamin Wu +5
Vision-Language Models (VLMs), such as CLIP, have shown strong zero-shot generalization but remain highly vulnerable to adversarial perturbations, posing serious risks in real-worl…
cs.CV2024★ 24 cited
Multi-modal Attribute Prompting for Vision-Language Models
Xin Liu, Jiamin Wu, and Wenfei Yang +2
Pre-trained Vision-Language Models (VLMs), like CLIP, exhibit strong generalization ability to downstream tasks but struggle in few-shot scenarios. Existing prompting techniques pr…