1 paper
Hwaran Lee, Seokhee Hong, Joonsuk Park +10
The potential social harms that large language models pose, such as generating offensive content and reinforcing biases, are steeply rising. Existing works focus on coping with thi…