1 paper
Théo Lasnier, Wissam Antoun, Francis Kulumba +2
Backdoor attacks pose significant security risks for Large Language Models (LLMs), yet the internal mechanisms by which triggers operate remain poorly understood. We present the fi…