7 citations · 8 across the 4 of their papers we have counts for
1 paper · 1 filter
Leila Khalatbari, Yejin Bang, Dan Su +4
Conversational models that are generative and open-domain are particularly susceptible to generating unsafe content since they are trained on web-based social data. Prior approache…