1 paper
Mohammad Akbar-Tajari, Mohammad Taher Pilehvar, Mohammad Mahmoody
The challenge of ensuring Large Language Models (LLMs) align with societal standards is of increasing interest, as these models are still prone to adversarial jailbreaks that bypas…