From the 1 of 8 linked papers with an AI index.
8 papers
The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction
Mirko Lai, Alessandra Urbinati, Simona Frenda +2
The paper proposes a framework that uses conformal prediction and collaborative‑filtering style annotator representations to study how large language models agree or disagree with…
SemEval-2026 Task 9: Detecting Multilingual, Multicultural and Multievent Online Polarization
Usman Naseem, Robert Geislinger, Juan Ren +31
We present SemEval-2026 Task 9, a shared task on online polarization detection, covering 22 languages and comprising over 110K annotated instances. Each data instance is multi-labe…
Epistemic Injustice in Language Models: An Audit of Pretraining Filters and Guardrails
Marco Antonio Stranisci, A Pranav, Rossana Damiano +2
Modern language models rely on pretraining filters to remove undesirable content from training corpora and inference-time guardrails to suppress undesirable outputs during deployme…
TailNLG: A Multilingual Benchmark Addressing Verbalization of Long-Tail Entities
Lia Draetta, Michael Oliverio, Virginia Ramón-Ferrer +6
The automatic verbalization of structured knowledge is a key task for making knowledge graphs accessible to non-expert users and supporting retrieval-augmented generation systems.…
Conspiracy Frame: a Semiotically-Driven Approach for Conspiracy Theories Detection
Heidi Campana Piva, Shaina Ashraf, Maziar Kianimoghadam Jouneghani +4
Conspiracy theories are anti-authoritarian narratives that lead to social conflict, impacting how people perceive political information. To help in understanding this issue, we int…
Are you sure? Measuring models bias in content moderation through uncertainty
Alessandra Urbinati, Mirko Lai, Simona Frenda +1
Automatic content moderation is crucial to ensuring safety in social media. Language Model-based classifiers are being increasingly adopted for this task, but it has been shown tha…