2 citations · 5 across the 19 of their papers we have counts for
1 paper · 1 filter
Jindong Li, Ying Liu, Yali Fu +4
LLMs are increasingly equipped with safety alignment mechanisms, yet recent studies demonstrate that they remain vulnerable to jailbreaking attacks that elicit harmful behaviors wi…