activity
20192022
most citedNADI 2020: The First Nuanced Arabic Dialect Identification Shared Task

57 citations · 224 across the 26 of their papers we have counts for

collaborators
Showing cs.CLShow all

27 papers · 1 filter

cs.CL2022

NADI 2022: The Third Nuanced Arabic Dialect Identification Shared Task

Muhammad Abdul-Mageed, Chiyu Zhang, AbdelRahim Elmadany +2

We describe findings of the third Nuanced Arabic Dialect Identification Shared Task (NADI 2022). NADI aims at advancing state of the art Arabic NLP, including on Arabic dialects. I…

cs.CL2022

Decay No More: A Persistent Twitter Dataset for Learning Social Meaning

Chiyu Zhang, Muhammad Abdul-Mageed, El Moatez Billah Nagoudi

With the proliferation of social media, many studies resort to social media to construct datasets for developing social meaning understanding systems. For the popular case of Twitt…

cs.CL20222 cited

Automatic Detection of Entity-Manipulated Text using Factual Knowledge

Ganesh Jawahar, Muhammad Abdul-Mageed, Laks V. S. Lakshmanan

In this work, we focus on the problem of distinguishing a human written news article from a news article that is created by manipulating entities in a human written news article (e…

cs.CL20222 cited

Towards Afrocentric NLP for African Languages: Where We Are and Where We Can Go

Ife Adebara, Muhammad Abdul-Mageed

Aligning with ACL 2022 special Theme on "Language Diversity: from Low Resource to Endangered Languages", we discuss the major linguistic and sociopolitical challenges facing develo…

cs.CL20213 cited

Machine Translation of Low-Resource Indo-European Languages

Wei-Rui Chen, Muhammad Abdul-Mageed

In this work, we investigate methods for the challenging task of translating between low-resource language pairs that exhibit some level of similarity. In particular, we consider t…

cs.CL20211 cited

Exploring Text-to-Text Transformers for English to Hinglish Machine Translation with Synthetic Code-Mixing

Ganesh Jawahar, El Moatez Billah Nagoudi, Muhammad Abdul-Mageed +1

We describe models focused at the understudied problem of translating between monolingual and code-mixed language pairs. More specifically, we offer a wide range of models that con…