1 paper
Arash Raftari, Mehrdad Mahdavi, Nathan Blackthorn +1
Backdoor attacks pose a serious threat to large language models (LLMs) by causing otherwise benign systems to produce attacker-specified malicious behavior when a hidden trigger is…