1 paper
Reinelle Jan Bugnot, Soohyeon Choi, Hoon Wei Lim +1
Jailbreaking attacks on large language models pose a significant threat to AI safety by enabling the generation of harmful or restricted content. While prior work has explored both…