1 paper · 1 filter
Genglin Liu, Muye Zhang, Krishnamurthy Viswanathan +5
Multimodal Large Language Models (MLLMs) are increasingly deployed for nuanced content safety and moderation tasks, yet they remain vulnerable to adversarial attacks and out-of-dis…