1 paper
Yunlang Dai, Emma Lurie, Danaé Metaxa +1
Large language models' (LLMs') outputs are shaped by opaque and frequently-changing company content moderation policies and practices. LLM moderation often takes the form of refusa…