1 paper · 1 filter
Viswonathan Manoranjan, Amogh Gupta, Anvesh Rao Vijjini +2
Large language models often struggle with sensitive prompts. They may refuse outright, provide generic safety boilerplate, or fail to address the user's legitimate informational ne…