2 papers
cs.LG2026
Measuring the Prevalence of Policy Violating Content with ML Assisted Sampling and LLM Labeling
Attila Dobi, Aravindh Manickavasagam, Benjamin Thompson +2
Content safety teams need metrics that reflect what users actually experience, not only what is reported. We study prevalence: the fraction of user views (impressions) that went to…
stat.AP2026
Decision Quality Evaluation Framework at Pinterest
Yuqi Tian, Robert Paine, Attila Dobi +3
Online platforms require robust systems to enforce content safety policies at scale. A critical component of these systems is the ability to evaluate the quality of moderation deci…