1 paper · 1 filter
Sarah Ball, Greg Gluch, Shafi Goldwasser +3
With the increased deployment of large language models (LLMs), one concern is their potential misuse for generating harmful content. Our work studies the alignment challenge, with…