1 paper
Kazuki Iwahana, Masaru Matsubayashi, Takuma Koyama +3
Backdoor attacks pose a serious threat to the safety and reliability of Large Language Models (LLMs), as they cause models to behave normally on clean inputs while producing attack…