337 citations · 543 across the 19 of their papers we have counts for
Showing 2020 · cs.CLShow all
2 papers · 2 filters
cs.CL2020
"Nice Try, Kiddo": Investigating Ad Hominems in Dialogue Responses
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan +1
Ad hominem attacks are those that target some feature of a person's character instead of the position the person is maintaining. These attacks are harmful because they propagate im…
cs.CL2020
Towards Controllable Biases in Language Generation
Emily Sheng, Kai-Wei Chang, Premkumar Natarajan +1
We present a general approach towards controllable societal biases in natural language generation (NLG). Building upon the idea of adversarial triggers, we develop a method to indu…