1 paper
Qin Yang, Lu Malloy, Joshua Lee +4
Large language model (LLM)-powered content moderation systems are a critical defense against harmful online content. However, they operate primarily on tokenized text and often ove…