22 citations · 34 across the 19 of their papers we have counts for
1 paper · 2 filters
Agam Goyal, Xianyang Zhan, Yilun Chen +2
Large language models (LLMs) have shown great potential in flagging harmful content in online communities. Yet, existing approaches for moderation require a separate model for ever…