1 paper · 1 filter
Tingxu Han, Wei Song, Ziqi Ding +6
Large language models (LLMs) increasingly mediate decisions in domains where unfair treatment of demographic groups is unacceptable. Existing work probes when biased outputs appear…