3 citations · 4 across the 10 of their papers we have counts for
13 papers · 1 filter
Konooz: Multi-domain Multi-dialect Corpus for Named Entity Recognition
Nagham Hamad, Mohammed Khalilia, Mustafa Jarrar
We introduce Konooz, a novel multi-dimensional corpus covering 16 Arabic dialects across 10 domains, resulting in 160 distinct corpora. The corpus comprises about 777k tokens, care…
Event-Arguments Extraction Corpus and Modeling using BERT for Arabic
Alaa Aljabari, Lina Duaibes, Mustafa Jarrar +1
Event-argument extraction is a challenging task, particularly in Arabic due to sparse linguistic resources. To fill this gap, we introduce the \hadath corpus (k tokens) as an…
ArabicNLU 2024: The First Arabic Natural Language Understanding Shared Task
Mohammed Khalilia, Sanad Malaysha, Reem Suwaileh +4
This paper presents an overview of the Arabic Natural Language Understanding (ArabicNLU 2024) shared task, focusing on two subtasks: Word Sense Disambiguation (WSD) and Location Me…
WojoodNER 2024: The Second Arabic Named Entity Recognition Shared Task
Mustafa Jarrar, Nagham Hamad, Mohammed Khalilia +3
We present WojoodNER-2024, the second Arabic Named Entity Recognition (NER) Shared Task. In WojoodNER-2024, we focus on fine-grained Arabic NER. We provided participants with a new…
AraFinNLP 2024: The First Arabic Financial NLP Shared Task
Sanad Malaysha, Mo El-Haj, Saad Ezzini +5
The expanding financial markets of the Arab world require sophisticated Arabic NLP tools. To address this need within the banking domain, the Arabic Financial NLP (AraFinNLP) share…
NLU-STR at SemEval-2024 Task 1: Generative-based Augmentation and Encoder-based Scoring for Semantic Textual Relatedness
Sanad Malaysha, Mustafa Jarrar, Mohammed Khalilia
Semantic textual relatedness is a broader concept of semantic similarity. It measures the extent to which two chunks of text convey similar meaning or topics, or share related conc…