collaborators

9 papers

cs.CL2026

Bulbul: A Dataset for Dialectal Arabic Speech Recognition

Ahmed Ashraf, Aisha Alansari, Fadel Al Abbas +30

Arabic automatic speech recognition (ASR) faces unique challenges due to diglossia, extensive regional dialect variation, and limited speech resources. Existing speech datasets oft…

cs.CL2026

Mawqif-XT: An Arabic Benchmark Dataset for Cross-Target Stance Detection

Rasha Albalawi, Nuha Albadi, Hamzah Luqman +4

Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This paper presents the Mawqif-XT,…

cs.CL2026

CrossHallu: Do Hallucination Signals Generalize Across Languages and Domains in Large Language Model's Internals?

Aisha Alansari, Malak Alkhorasani, Hamzah Luqman

Recent hallucination detection techniques in large language models (LLMs) focus on directly extracting features from a model's internal representations and training a classifier on…

cs.CL2026

HalluScore: Large Language Model Hallucination Question Answering Benchmark

Aisha Alansari, Hamzah Luqman

Large language models (LLMs) have achieved remarkable progress in natural language generation, but remain susceptible to hallucination. In response to growing concerns about halluc…

cs.CL2026

Large Language Models Hallucination: A Comprehensive Survey

Aisha Alansari, Hamzah Luqman

Large language models (LLMs) have transformed natural language processing, achieving remarkable performance across diverse tasks. However, their impressive fluency often comes at t…

cs.CL2025

AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP

Ahmed Hasanaath, Aisha Alansari, Ahmed Ashraf +3

Large language models (LLMs) have shown remarkable progress in reasoning abilities and general natural language processing (NLP) tasks, yet their performance on Arabic data, charac…