1 paper
Jiahui Li, Yongchang Hao, Haoyu Xu +2
Despite the advancements in training Large Language Models (LLMs) with alignment techniques to enhance the safety of generated content, these models remain susceptible to jailbreak…