1 citations · 1 across the 1 of their papers we have counts for
1 paper
Shenyi Zhang, Yuchen Zhai, Keyan Guo +7
Despite the implementation of safety alignment strategies, large language models (LLMs) remain vulnerable to jailbreak attacks, which undermine these safety guardrails and pose sig…