30 citations · 33 across the 9 of their papers we have counts for
1 paper · 2 filters
Jinhao Pan, Chahat Raj, Anjishnu Mukherjee +4
Large language models (LLMs) exhibit social biases that reinforce harmful stereotypes, limiting their safe deployment. Most existing debiasing methods adopt a suppressive paradigm…