9 papers
Bulbul: A Dataset for Dialectal Arabic Speech Recognition
Ahmed Ashraf, Aisha Alansari, Fadel Al Abbas +30
Arabic automatic speech recognition (ASR) faces unique challenges due to diglossia, extensive regional dialect variation, and limited speech resources. Existing speech datasets oft…
Mawqif-XT: An Arabic Benchmark Dataset for Cross-Target Stance Detection
Rasha Albalawi, Nuha Albadi, Hamzah Luqman +4
Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This paper presents the Mawqif-XT,…
CrossHallu: Do Hallucination Signals Generalize Across Languages and Domains in Large Language Model's Internals?
Aisha Alansari, Malak Alkhorasani, Hamzah Luqman
Recent hallucination detection techniques in large language models (LLMs) focus on directly extracting features from a model's internal representations and training a classifier on…
HalluScore: Large Language Model Hallucination Question Answering Benchmark
Aisha Alansari, Hamzah Luqman
Large language models (LLMs) have achieved remarkable progress in natural language generation, but remain susceptible to hallucination. In response to growing concerns about halluc…
Large Language Models Hallucination: A Comprehensive Survey
Aisha Alansari, Hamzah Luqman
Large language models (LLMs) have transformed natural language processing, achieving remarkable performance across diverse tasks. However, their impressive fluency often comes at t…
AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP
Ahmed Hasanaath, Aisha Alansari, Ahmed Ashraf +3
Large language models (LLMs) have shown remarkable progress in reasoning abilities and general natural language processing (NLP) tasks, yet their performance on Arabic data, charac…