Showing cs.CLShow all
2 papers · 1 filter
cs.CL2025
Longitudinal Monitoring of LLM Content Moderation of Social Issues
Yunlang Dai, Emma Lurie, Danaé Metaxa +1
Large language models' (LLMs') outputs are shaped by opaque and frequently-changing company content moderation policies and practices. LLM moderation often takes the form of refusa…
cs.CL2025
Identity-related Speech Suppression in Generative AI Content Moderation
Grace Proebsting, Oghenefejiro Isaacs Anigboro, Charlie M. Crawford +2
Automated content moderation has long been used to help identify and filter undesired user-generated content online. But such systems have a history of incorrectly flagging content…