1 paper · 1 filter
Ziyu Liu, Tao Li, Tianjie Ni +5
Data-poisoning backdoors pose a practical threat to the fine-tuning of large language models (LLMs). Most existing attacks bind an attacker-selected behavior to fixed tokens, phras…