2 citations · 2 across the 1 of their papers we have counts for
1 paper
Kyra Yee, Alice Schoenauer Sebag, Olivia Redfield +3
Harmful content detection models tend to have higher false positive rates for content from marginalized groups. In the context of marginal abuse modeling on Twitter, such dispropor…