Showing cs.CLShow all
3 papers · 1 filter
cs.CL2024
Verifying the Robustness of Automatic Credibility Assessment
Piotr PrzybyÅa, Alexander Shvets, Horacio Saggion
Text classification methods have been widely investigated as a way to detect content of low credibility: fake news, social media bots, propaganda, etc. Quite accurate models (likel…
cs.CL2024
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
Yiping Jin, Leo Wanner, Alexander Shvets
Online hate detection suffers from biases incurred in data sampling, annotation, and model pre-training. Therefore, measuring the averaged performance over all examples in held-out…
cs.CL2024
Towards Weakly-Supervised Hate Speech Classification Across Datasets
Yiping Jin, Leo Wanner, Vishakha Laxman Kadam +1
As pointed out by several scholars, current research on hate speech (HS) recognition is characterized by unsystematic data creation strategies and diverging annotation schemata. Su…