Showing cs.CVShow all
2 papers · 1 filter
cs.CV2023
Interpreting and Controlling Vision Foundation Models via Text Explanations
Haozhe Chen, Junfeng Yang, Carl Vondrick +1
Large-scale pre-trained vision foundation models, such as CLIP, have become de facto backbones for various vision tasks. However, due to their black-box nature, understanding the u…
cs.CV2023
Test-time Detection and Repair of Adversarial Samples via Masked Autoencoder
Yun-Yun Tsai, Ju-Chin Chao, Albert Wen +4
Training-time defenses, known as adversarial training, incur high training costs and do not generalize to unseen attacks. Test-time defenses solve these issues but most existing te…