1 paper
Yi Zhao, Youzhi Zhang
Large language models (LLMs) are widely used in real-world applications, raising concerns about their safety and trustworthiness. While red-teaming with jailbreak prompts exposes t…