2 papers
cs.CV2025
Watch, Listen, Understand, Mislead: Tri-modal Adversarial Attacks on Short Videos for Content Appropriateness Evaluation
Sahid Hossain Mustakim, S M Jishanul Islam, Ummay Maria Muna +6
Multimodal Large Language Models (MLLMs) are increasingly used for content moderation, yet their robustness in short-form video contexts remains underexplored. Current safety evalu…
cs.CV2024
MIMIC: Multimodal Islamophobic Meme Identification and Classification
S M Jishanul Islam, Sahid Hossain Mustakim, Sadia Ahmmed +4
Anti-Muslim hate speech has emerged within memes, characterized by context-dependent and rhetorical messages using text and images that seemingly mimic humor but convey Islamophobi…