1 paper
Molly Apsel, Michael N. Jones
Drawing on constructs from psychology, prior work has identified a distinction between explicit and implicit bias in large language models (LLMs). While many LLMs undergo post-trai…