collaborators

18 papers

cs.CL2026

Romanized Arabic Across Dialects: Views, Usage Patterns, and Linguistic Variation

Amr Keleg, Ahmed Amine Ben Abdallah, Taha Yassine +3

Arabizi refers to Arabic written in Latin script. Although previous studies have shown that the prevalence and usage of Arabizi vary by factors such as region and age group, most N…

cs.CL2026

MEDIAREF: A Public Knowledge Store for Media Background Checks

Benjamin Nichols, Michael Schlichtkrull, Nedjma Ousidhoum

LLM-based retrieval-augmented generation (RAG) is increasingly used for automated fact-checking (AFC) and related tasks. By grounding LLM outputs in retrieved evidence, RAG-based s…

cs.CL2026

Algorithmic Fragility and Persona Bias in LLM-Generated Autistic Communication

Naba Rizvi, Mohammed Rizvi, Harper Strickland +2

Safety alignment reduces explicitly harmful outputs but inadvertently encodes a sanitized, neuronormative representation of marginalized communication. We investigate this encoding…

cs.CL2026

Text Analytics Evaluation Framework: A Case Study on LLMs and Social Media

Yuefeng Shi, Nedjma Ousidhoum, Jose Camacho-Collados

LLMs have demonstrated exceptional proficiency in a wide range of NLP tasks. However, a notable gap remains in practical data analysis scenarios, particularly when LLMs are require…

cs.CL2026

SemEval-2026 Task 7: Everyday Knowledge Across Diverse Languages and Cultures

Nedjma Ousidhoum, Junho Myung, Carla Perez-Almendros +27

We present our shared task on evaluating the adaptability of LLMs and NLP systems across multiple languages and cultures. The task data consist of an extended version of our manual…

cs.CL2026

Annotating Dimensions of Social Perception in Text: A Sentence-Level Dataset of Warmth and Competence

Mutaz Ayesh, Saif M. Mohammad, Nedjma Ousidhoum

Warmth (W) (often further broken down intoTrust (T) and Sociability (S)) and Competence (C) are central dimensions along which people evaluate individuals and social groups (Fiske,…