Showing cs.LGShow all
2 papers · 1 filter
cs.LG2022
Bandits for Online Calibration: An Application to Content Moderation on Social Media Platforms
Vashist Avadhanula, Omar Abdul Baki, Hamsa Bastani +17
We describe the current content moderation strategy employed by Meta to remove policy-violating content from its platforms. Meta relies on both handcrafted and learned risk models…
cs.LG2020
Multi-armed Bandits with Cost Subsidy
Deeksha Sinha, Karthik Abinav Sankararama, Abbas Kazerouni +1
In this paper, we consider a novel variant of the multi-armed bandit (MAB) problem, MAB with cost subsidy, which models many real-life applications where the learning agent has to…