2 citations · 2 across the 4 of their papers we have counts for
4 papers
ArBanking77: Intent Detection Neural Model and a New Dataset in Modern and Dialectical Arabic
Mustafa Jarrar, Ahmet Birim, Mohammed Khalilia +2
This paper presents the ArBanking77, a large Arabic dataset for intent detection in the banking domain. Our dataset was arabized and localized from the original English Banking77 d…
SALMA: Arabic Sense-Annotated Corpus and WSD Benchmarks
Mustafa Jarrar, Sanad Malaysha, Tymaa Hammouda +1
SALMA, the first Arabic sense-annotated corpus, consists of ~34K tokens, which are all sense-annotated. The corpus is annotated using two different sense inventories simultaneously…
WojoodNER 2023: The First Arabic Named Entity Recognition Shared Task
Mustafa Jarrar, Muhammad Abdul-Mageed, Mohammed Khalilia +4
We present WojoodNER-2023, the first Arabic Named Entity Recognition (NER) Shared Task. The primary focus of WojoodNER-2023 is on Arabic NER, offering novel NER datasets (i.e., Woj…
Context-Gloss Augmentation for Improving Arabic Target Sense Verification
Sanad Malaysha, Mustafa Jarrar, Mohammed Khalilia
Arabic language lacks semantic datasets and sense inventories. The most common semantically-labeled dataset for Arabic is the ArabGlossBERT, a relatively small dataset that consist…