Showing cs.CVShow all
2 papers · 1 filter
cs.CV2024
AEMIM: Adversarial Examples Meet Masked Image Modeling
Wenzhao Xiang, Chang Liu, Hang Su +1
Masked image modeling (MIM) has gained significant traction for its remarkable prowess in representation learning. As an alternative to the traditional approach, the reconstruction…
cs.CV2023
Machine Vision Therapy: Multimodal Large Language Models Can Enhance Visual Robustness via Denoising In-Context Learning
Zhuo Huang, Chang Liu, Yinpeng Dong +3
Although vision models such as Contrastive Language-Image Pre-Training (CLIP) show impressive generalization performance, their zero-shot robustness is still limited under Out-of-D…