1 paper · 1 filter
Xiao Li, Wei Zhang, Zhuhong Li +6
Aligned Large Language Models (LLMs) have attracted significant attention for their safety, particularly in the context of jailbreak attacks that attempt to bypass guardrails via a…