1 paper
Yizhe Zeng, Chenxu Niu, Wei Zhang +7
Backdoor attacks pose a serious threat to large language models (LLMs), but existing defenses remain fragmented, failing to pro?vide unified defense against both dirty-label and cl…