1 citations · 2 across the 4 of their papers we have counts for
4 papers
Don't sweat the small stuff, classify the rest: Sample Shielding to protect text classifiers against adversarial attacks
Jonathan Rusert, Padmini Srinivasan
Deep learning (DL) is being used extensively for text classification. However, researchers have demonstrated the vulnerability of such classifiers to adversarial attacks. Attackers…
A Girl Has A Name, And It's ... Adversarial Authorship Attribution for Deobfuscation
Wanyue Zhai, Jonathan Rusert, Zubair Shafiq +1
Recent advances in natural language processing have enabled powerful privacy-invasive authorship attribution. To counter authorship attribution, researchers have proposed a variety…
Suum Cuique: Studying Bias in Taboo Detection with a Community Perspective
Osama Khalid, Jonathan Rusert, Padmini Srinivasan
Prior research has discussed and illustrated the need to consider linguistic norms at the community level when studying taboo (hateful/offensive/toxic etc.) language. However, a me…
On The Robustness of Offensive Language Classifiers
Jonathan Rusert, Zubair Shafiq, Padmini Srinivasan
Social media platforms are deploying machine learning based offensive language classification systems to combat hateful, racist, and other forms of offensive speech at scale. Howev…