4 papers
Disentangling Hate Across Target Identities
Yiping Jin, Leo Wanner, Aneesh Moideen Koya
Hate speech (HS) classifiers do not perform equally well in detecting hateful expressions towards different target identities. They also demonstrate systematic biases in predicted…
ARAIDA: Analogical Reasoning-Augmented Interactive Data Annotation
Chen Huang, Yiping Jin, Ilija Ilievski +2
Human annotation is a time-consuming task that requires a significant amount of effort. To address this issue, interactive data annotation utilizes an annotation model to provide s…
GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?
Yiping Jin, Leo Wanner, Alexander Shvets
Online hate detection suffers from biases incurred in data sampling, annotation, and model pre-training. Therefore, measuring the averaged performance over all examples in held-out…
Towards Weakly-Supervised Hate Speech Classification Across Datasets
Yiping Jin, Leo Wanner, Vishakha Laxman Kadam +1
As pointed out by several scholars, current research on hate speech (HS) recognition is characterized by unsystematic data creation strategies and diverging annotation schemata. Su…