65 citations · 160 across the 31 of their papers we have counts for
3 papers · 1 filter
Learning from the Worst: Dynamically Generated Datasets to Improve Online Hate Detection
Bertie Vidgen, Tristan Thrush, Zeerak Waseem +1
We present a human-and-model-in-the-loop process for dynamically generating datasets and training better performing and more robust hate detection models. We provide a new dataset…
HateCheck: Functional Tests for Hate Speech Detection Models
Paul Röttger, Bertram Vidgen, Dong Nguyen +3
Detecting online hate is a difficult task that even state-of-the-art models struggle with. Typically, hate speech detection models are evaluated by measuring their performance on h…
Detecting East Asian Prejudice on Social Media
Bertie Vidgen, Austin Botelho, David Broniatowski +6
The outbreak of COVID-19 has transformed societies across the world as governments tackle the health, economic and social costs of the pandemic. It has also raised concerns about t…