1 paper · 1 filter
Wonjun Lee, Doehyeon Lee, Eugene Choi +5
Current Vision Language Models (VLMs) remain vulnerable to malicious prompts that induce harmful outputs. Existing safety benchmarks for VLMs primarily rely on automated evaluation…