2 citations · 6 across the 8 of their papers we have counts for
8 papers
Arab Voices: Mapping Standard and Dialectal Arabic Speech Technology
Peter Sullivan, AbdelRahim Elmadany, Alcides Alcoba Inciarte +1
Dialectal Arabic (DA) speech data vary widely in domain coverage, dialect labeling practices, and recording conditions, complicating cross-dataset comparison and model evaluation.…
Pearl: A Multimodal Culturally-Aware Arabic Instruction Dataset
Fakhraddin Alwajih, Samar M. Magdy, Abdellah El Mekki +34
Mainstream large vision-language models (LVLMs) inherently encode cultural biases, highlighting the need for diverse multimodal datasets. To address this gap, we introduce PEARL, a…
Voice of a Continent: Mapping Africa's Speech Technology Frontier
AbdelRahim Elmadany, Sang Yun Kwon, Hawau Olamide Toyin +3
Africa's rich linguistic diversity remains significantly underrepresented in speech technologies, creating barriers to digital inclusion. To alleviate this challenge, we systematic…
Violet: A Vision-Language Model for Arabic Image Captioning with Gemini Decoder
Abdelrahman Mohamed, Fakhraddin Alwajih, El Moatez Billah Nagoudi +2
Although image captioning has a vast array of applications, it has not reached its full potential in languages other than English. Arabic, for instance, although the native languag…
Zero-Shot Slot and Intent Detection in Low-Resource Languages
Sang Yun Kwon, Gagan Bhatia, El Moatez Billah Nagoudi +2
Intent detection and slot filling are critical tasks in spoken and natural language understanding for task-oriented dialog systems. In this work we describe our participation in th…
SERENGETI: Massively Multilingual Language Models for Africa
Ife Adebara, AbdelRahim Elmadany, Muhammad Abdul-Mageed +1
Multilingual pretrained language models (mPLMs) acquire valuable, generalizable linguistic information during pretraining and have advanced the state of the art on task-specific fi…