1 citations · 1 across the 1 of their papers we have counts for
1 paper
Shuo Yang, Qihui Zhang, Yuyang Liu +7
Fine-tuning large language models (LLMs) improves performance but introduces critical safety vulnerabilities: even minimal harmful data can severely compromise safety measures. We…