2 citations · 5 across the 12 of their papers we have counts for
12 papers
GenAI Content Detection Task 1: English and Multilingual Machine-Generated Text Detection: AI vs. Human
Yuxia Wang, Artem Shelmanov, Jonibek Mansurov +23
We present the GenAI Content Detection Task~1 -- a shared task on binary machine generated text detection, conducted as a part of the GenAI workshop at COLING 2025. The task consis…
The FIGNEWS Shared Task on News Media Narratives
Wajdi Zaghouani, Mustafa Jarrar, Nizar Habash +5
We present an overview of the FIGNEWS shared task, organized as part of the ArabicNLP 2024 conference co-located with ACL 2024. The shared task addresses bias and propaganda annota…
NADI 2024: The Fifth Nuanced Arabic Dialect Identification Shared Task
Muhammad Abdul-Mageed, Amr Keleg, AbdelRahim Elmadany +5
We describe the findings of the fifth Nuanced Arabic Dialect Identification Shared Task (NADI 2024). NADI's objective is to help advance SoTA Arabic NLP by providing guidance, data…
Strategies for Arabic Readability Modeling
Juan Piñeros Liberato, Bashar Alhafni, Muhamed Al Khalil +1
Automatic readability assessment is relevant to building NLP applications for education, content analysis, and accessibility. However, Arabic readability assessment is a challengin…
Arabic Diacritics in the Wild: Exploiting Opportunities for Improved Diacritization
Salman Elgamal, Ossama Obeid, Tameem Kabbani +2
The widespread absence of diacritical marks in Arabic text poses a significant challenge for Arabic natural language processing (NLP). This paper explores instances of naturally oc…
The SAMER Arabic Text Simplification Corpus
Bashar Alhafni, Reem Hazim, Juan Piñeros Liberato +2
We present the SAMER Corpus, the first manually annotated Arabic parallel corpus for text simplification targeting school-aged learners. Our corpus comprises texts of 159K words se…