1 paper · 1 filter
Xuechunzi Bai, Angelina Wang, Ilia Sucholutsky +1
Large language models (LLMs) can pass explicit social bias tests but still harbor implicit biases, similar to humans who endorse egalitarian beliefs yet exhibit subtle biases. Meas…