1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.CR2025
Defending against Backdoor Attacks via Module Switching
Weijun Li, Ansh Arora, Xuanli He +2
Backdoor attacks pose a serious threat to deep neural networks (DNNs), allowing adversaries to implant triggers for hidden behaviors in inference. Defending against such vulnerabil…
cs.CL2024★ 1 cited
Here's a Free Lunch: Sanitizing Backdoored Models with Model Merge
Ansh Arora, Xuanli He, Maximilian Mozes +3
The democratization of pre-trained language models through open-source initiatives has rapidly advanced innovation and expanded access to cutting-edge technologies. However, this o…