5 citations · 5 across the 3 of their papers we have counts for
3 papers
Towards a Unified Benchmark for Arabic Pronunciation Assessment: Quranic Recitation as Case Study
Yassine El Kheir, Omnia Ibrahim, Amit Meghanani +12
We present a unified benchmark for mispronunciation detection in Modern Standard Arabic (MSA) using Qur'anic recitation as a case study. Our approach lays the groundwork for advanc…
MorphBPE: A Morpho-Aware Tokenizer Bridging Linguistic Complexity for Efficient LLM Training Across Morphologies
Ehsaneddin Asgari, Yassine El Kheir, Mohammad Ali Sadraei Javaheri
Tokenization is fundamental to Natural Language Processing (NLP), directly impacting model efficiency and linguistic fidelity. While Byte Pair Encoding (BPE) is widely used in Larg…
Fanar: An Arabic-Centric Multimodal Generative AI Platform
Fanar Team, Ummar Abbas, Mohammad Shahmeer Ahmad +39
We present Fanar, a platform for Arabic-centric multimodal generative AI systems, that supports language, speech and image generation tasks. At the heart of Fanar are Fanar Star an…