1 paper · 1 filter
Soorya Ram Shimgekar, Agam Goyal, Amruta Parulekar +6
Large language models (LLMs) are increasingly deployed in conversational settings where user tone ranges from polite to adversarial or toxic, yet less is known about whether toxic…