1 paper · 1 filter
Xurui Song, Zhixin Xie, Shuo Huai +2
The wide adoption of Large Language Models (LLMs) has attracted significant attention from jailbreak attacks, where adversarial prompts crafted through optimization or m…