3 papers
cs.CV2026
A Controlled Study of Self-Supervised Image and Video Pretraining under Limited Resources
Brunó B. Englert, Gijs Dubbelman
Visual foundation models are a cornerstone of image and video understanding but typically require large amounts of data and computation. The current scale required for pretraining…
cs.CL2025
Exploring the Limits of Zero Shot Vision Language Models for Hate Meme Detection: The Vulnerabilities and their Interpretations
Naquee Rizwan, Paramananda Bhaskar, Mithun Das +3
There is a rapid increase in the use of multimedia content in current social media platforms. One of the highly popular forms of such multimedia content are memes. While memes have…
cs.CL2024
CrowdCounter: A benchmark type-specific multi-target counterspeech dataset
Punyajoy Saha, Abhilash Datta, Abhik Jana +1
Counterspeech presents a viable alternative to banning or suspending users for hate speech while upholding freedom of expression. However, writing effective counterspeech is challe…