6 citations · 7 across the 6 of their papers we have counts for
1 paper · 1 filter
Yongjian Guo, Wanlun Ma, Lingyu Shen +2
Fine-tuning is the dominant paradigm for specializing large language models (LLMs), yet it exposes a critical vulnerability: malicious data providers can embed harmful behaviors in…