1 paper · 1 filter
Xiaohu Li, Yunfeng Ning, Zepeng Bao +3
Security alignment enables the Large Language Model (LLM) to gain the protection against malicious queries, but various jailbreak attack methods reveal the vulnerability of this se…