1 citations · 1 across the 2 of their papers we have counts for
2 papers
cs.AI2025★ 1 cited
Scaling Reinforcement Learning for Content Moderation with Large Language Models
Hamed Firooz, Rui Liu, Yuchen Lu +15
Content moderation at scale remains one of the most pressing challenges in today's digital ecosystem, where billions of user- and AI-generated artifacts must be continuously evalua…
cs.LG2022
Bandits for Online Calibration: An Application to Content Moderation on Social Media Platforms
Vashist Avadhanula, Omar Abdul Baki, Hamsa Bastani +17
We describe the current content moderation strategy employed by Meta to remove policy-violating content from its platforms. Meta relies on both handcrafted and learned risk models…