2 papers
cs.CL2026
The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction
Mirko Lai, Alessandra Urbinati, Simona Frenda +2
Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in settings where LLMs are incr…
cs.CL2025
Are you sure? Measuring models bias in content moderation through uncertainty
Alessandra Urbinati, Mirko Lai, Simona Frenda +1
Automatic content moderation is crucial to ensuring safety in social media. Language Model-based classifiers are being increasingly adopted for this task, but it has been shown tha…