7 citations · 7 across the 2 of their papers we have counts for
2 papers
cs.CL2025
MSTS: A Multimodal Safety Test Suite for Vision-Language Models
Paul Röttger, Giuseppe Attanasio, Felix Friedrich +19
Vision-language models (VLMs), which process image and text inputs, are increasingly integrated into chat assistants and other consumer AI applications. Without proper safeguards,…
cs.CL2022★ 7 cited
Hypothesis Engineering for Zero-Shot Hate Speech Detection
Janis Goldzycher, Gerold Schneider
Standard approaches to hate speech detection rely on sufficient available hate speech annotations. Extending previous work that repurposes natural language inference (NLI) models f…