1 paper · 1 filter
Adeel Yousaf, Joseph Fioresi, James Beetham +2
Improving the safety of vision-language models like CLIP via fine-tuning often comes at a steep price, causing significant drops in their generalization performance. We find this t…