1 paper · 1 filter
Haoming Yang, Ke Ma, Xiaojun Jia +3
Despite the remarkable performance of Large Language Models (LLMs), they remain vulnerable to jailbreak attacks, which can compromise their safety mechanisms. Existing studies ofte…