Showing cs.CLShow all
2 papers · 1 filter
cs.CL2020
A little goes a long way: Improving toxic language classification despite data scarcity
Mika Juuti, Tommi Gröndahl, Adrian Flanagan +1
Detection of some types of toxic language is hampered by extreme scarcity of labeled training data. Data augmentation - generating new synthetic data from a labeled seed dataset -…
cs.CL2018
All You Need is "Love": Evading Hate-speech Detection
Tommi Gröndahl, Luca Pajola, Mika Juuti +2
With the spread of social networks and their unfortunate use for hate speech, automatic detection of the latter has become a pressing problem. In this paper, we reproduce seven sta…