activity
20162025
most citedNADI 2024: The Fifth Nuanced Arabic Dialect Identification Shared Task

2 citations · 5 across the 12 of their papers we have counts for

collaborators

12 papers

cs.CL2025

GenAI Content Detection Task 1: English and Multilingual Machine-Generated Text Detection: AI vs. Human

Yuxia Wang, Artem Shelmanov, Jonibek Mansurov +23

We present the GenAI Content Detection Task~1 -- a shared task on binary machine generated text detection, conducted as a part of the GenAI workshop at COLING 2025. The task consis…

cs.CL2024

The FIGNEWS Shared Task on News Media Narratives

Wajdi Zaghouani, Mustafa Jarrar, Nizar Habash +5

We present an overview of the FIGNEWS shared task, organized as part of the ArabicNLP 2024 conference co-located with ACL 2024. The shared task addresses bias and propaganda annota…

cs.CL20242 cited

NADI 2024: The Fifth Nuanced Arabic Dialect Identification Shared Task

Muhammad Abdul-Mageed, Amr Keleg, AbdelRahim Elmadany +5

We describe the findings of the fifth Nuanced Arabic Dialect Identification Shared Task (NADI 2024). NADI's objective is to help advance SoTA Arabic NLP by providing guidance, data…

cs.CL2024

Strategies for Arabic Readability Modeling

Juan Piñeros Liberato, Bashar Alhafni, Muhamed Al Khalil +1

Automatic readability assessment is relevant to building NLP applications for education, content analysis, and accessibility. However, Arabic readability assessment is a challengin…

cs.CL2024

Arabic Diacritics in the Wild: Exploiting Opportunities for Improved Diacritization

Salman Elgamal, Ossama Obeid, Tameem Kabbani +2

The widespread absence of diacritical marks in Arabic text poses a significant challenge for Arabic natural language processing (NLP). This paper explores instances of naturally oc…

cs.CL2024

The SAMER Arabic Text Simplification Corpus

Bashar Alhafni, Reem Hazim, Juan Piñeros Liberato +2

We present the SAMER Corpus, the first manually annotated Arabic parallel corpus for text simplification targeting school-aged learners. Our corpus comprises texts of 159K words se…