1 paper · 1 filter
Miao Yu, Zhenhong Zhou, Moayad Aloqaily +5
Fine-tuned Large Language Models (LLMs) are vulnerable to backdoor attacks through data poisoning, yet the internal mechanisms governing these attacks remain a black box. Previous…