1 paper · 1 filter
Yongjian Guo, Wanlun Ma, Lingyu Shen +2
Fine-tuning is the dominant paradigm for specializing large language models (LLMs), yet it exposes a critical vulnerability: malicious data providers can embed harmful behaviors in…