1 paper
Wenpeng Xing, Moran Fang, Guangtai Wang +2
While Large Language Models (LLMs) have achieved remarkable performance, they remain vulnerable to jailbreak attacks that circumvent safety constraints. Existing strategies, rangin…