2 papers
cs.CL2023
Contextualizing the Limits of Model & Evaluation Dataset Curation on Semantic Similarity Classification Tasks
Daniel Theron
This paper demonstrates how the limitations of pre-trained models and open evaluation datasets factor into assessing the performance of binary semantic similarity classification ta…
stat.AP2023
Statistical Methods for Auditing the Quality of Manual Content Reviews
Xuan Yang, Andrew J Smart, Daniel Theron
Large technology firms face the problem of moderating content on their online platforms for compliance with laws and policies. To accomplish this at the scale of billions of pieces…