1 paper · 1 filter
Ayana Hussain, Patrick Zhao, Nicholas Vincent
Large Language Models (LLMs) are a double-edged sword capable of generating harmful misinformation -- inadvertently, or when prompted by "jailbreak" attacks that attempt to produce…