2 citations · 4 across the 3 of their papers we have counts for
1 paper · 1 filter
Kyra Yee, Alice Schoenauer Sebag, Olivia Redfield +3
Harmful content detection models tend to have higher false positive rates for content from marginalized groups. In the context of marginal abuse modeling on Twitter, such dispropor…