1 paper
Tong Zhang, Zexin Li, Simin Chen +1
Jailbreak defenses are essential for protecting large language models (LLMs), but they can also introduce secondary costs that weaken model utility. We present a systematic study o…