1 citations · 1 across the 12 of their papers we have counts for
1 paper · 1 filter
Leheng Sheng, Changshuo Shen, Weixiang Zhao +6
As LLMs are increasingly deployed in real-world applications, ensuring their ability to refuse malicious prompts, especially jailbreak attacks, is essential for safe and reliable u…