1 paper · 1 filter
Lang Gao, Jiahui Geng, Xiangliang Zhang +2
Jailbreaking in Large Language Models (LLMs) is a major security concern as it can deceive LLMs to generate harmful text. Yet, there is still insufficient understanding of how jail…