1 paper
Haibo Jin, Ruoxi Chen, Peiyan Zhang +2
The discovery of "jailbreaks" to bypass safety filters of Large Language Models (LLMs) and harmful responses have encouraged the community to implement safety measures. One major s…