5 citations · 6 across the 5 of their papers we have counts for
1 paper · 1 filter
Zhanhui Zhou, Lingjie Chen, Chao Yang +1
One way to mitigate risks in vision-language models (VLMs) is to remove dangerous samples in their training data. However, such data moderation can be easily bypassed when harmful…