6 citations · 6 across the 3 of their papers we have counts for
3 papers
A Constraint-Enforcing Reward for Adversarial Attacks on Text Classifiers
Tom Roth, Inigo Jauregi Unanue, Alsharif Abuadbba +1
Text classifiers are vulnerable to adversarial examples -- correctly-classified examples that are deliberately transformed to be misclassified while satisfying acceptability constr…
A Generative Adversarial Attack for Multilingual Text Classifiers
Tom Roth, Inigo Jauregi Unanue, Alsharif Abuadbba +1
Current adversarial attack algorithms, where an adversary changes a text to fool a victim model, have been repeatedly shown to be effective against text classifiers. These attacks,…
Token-Modification Adversarial Attacks for Natural Language Processing: A Survey
Tom Roth, Yansong Gao, Alsharif Abuadbba +2
Many adversarial attacks target natural language processing systems, most of which succeed through modifying the individual tokens of a document. Despite the apparent uniqueness of…