2 papers
cs.CV2025
Revealing Temporal Label Noise in Multimodal Hateful Video Classification
Shuonan Yang, Tailin Chen, Rahul Singh +3
The rapid proliferation of online multimedia content has intensified the spread of hate speech, presenting critical societal and regulatory challenges. While recent work has advanc…
cs.CV2023
Leveraging Generative Language Models for Weakly Supervised Sentence Component Analysis in Video-Language Joint Learning
Zaber Ibn Abdul Hakim, Najibul Haque Sarker, Rahul Pratap Singh +3
A thorough comprehension of textual data is a fundamental element in multi-modal video analysis tasks. However, recent works have shown that the current models do not achieve a com…