1 paper · 1 filter
Yanhui Li, Qi Zhou, Zhihong Xu +3
Large vision-language models (LVLMs) are increasingly used for tasks where detecting multimodal harmful content is crucial, such as online content moderation. However, real-world h…