1 paper · 1 filter
Dmitriy Bespalov, Sourav Bhabesh, Yi Xiang +2
Recent NLP literature pays little attention to the robustness of toxicity language predictors, while these systems are most likely to be used in adversarial contexts. This paper pr…