2 citations
1 paper
Guorui Chen, Yifan Xia, Xiaojun Jia +3
Large language models (LLMs) enhance security through alignment when widely used, but remain susceptible to jailbreak attacks capable of producing inappropriate content. Jailbreak…