6 citations · 9 across the 3 of their papers we have counts for
1 paper · 1 filter
Siwon Kim, Shuyang Dai, Mohammad Kachuee +3
Current conversational AI systems based on large language models (LLMs) are known to generate unsafe responses, agreeing to offensive user input or including toxic content. Previou…