241 citations · 272 across the 11 of their papers we have counts for
Showing 2023 · cs.LGShow all
2 papers · 2 filters
cs.LG2023★ 3 cited
Large Language Models Are Better Adversaries: Exploring Generative Clean-Label Backdoor Attacks Against Text Classifiers
Wencong You, Zayd Hammoudeh, Daniel Lowd
Backdoor attacks manipulate model predictions by inserting innocuous triggers into training and test data. We focus on more realistic and more challenging clean-label attacks where…
cs.LG2023
Provable Robustness Against a Union of Adversarial Attacks
Zayd Hammoudeh, Daniel Lowd
Sparse or adversarial attacks arbitrarily perturb an unknown subset of the features. robustness analysis is particularly well-suited for heterogeneous (tabular) d…