11 citations · 11 across the 2 of their papers we have counts for
2 papers
cs.LG2023
Characterizing the Optimal 0-1 Loss for Multi-class Classification with a Test-time Attacker
Sihui Dai, Wenxin Ding, Arjun Nitin Bhagoji +4
Finding classifiers robust to adversarial examples is critical for their safe deployment. Determining the robustness of the best possible classifier under a given threat model for…
cs.CR2022★ 11 cited
Post-breach Recovery: Protection against White-box Adversarial Examples for Leaked DNN Models
Shawn Shan, Wenxin Ding, Emily Wenger +2
Server breaches are an unfortunate reality on today's Internet. In the context of deep neural network (DNN) models, they are particularly harmful, because a leaked model gives an a…