11 papers
Aslema at NADI 2026: Augmentation through Fewshot for SLU
Tajwaar Shafiq, Hunzalah Hassan Bhatti, Shammur Absar Chowdhury +1
We present Aslema, our system for NADI 2026 Shared Task 5, which consists of two subtasks: intent recognition and slot filling. We evaluate four omni LLMs in a zero-shot setting an…
Multi-Task Instruction Tuning via Data Scheduling for Low-Resource Arabic SpeechLLMs
Hunzalah Hassan Bhatti, Firoj Alam, Shammur Absar Chowdhury
Audio large language models (LLMs) enable unified speech understanding and generation, but adapting them to linguistically complex and dialect-rich settings such as Arabic-English…
OASIS: A Multilingual and Multimodal Dataset for Culturally Grounded Spoken Visual QA
Firoj Alam, Ali Ezzat Shahroor, Md. Arid Hasan +8
Large-scale multimodal models achieve strong results on tasks like Visual Question Answering (VQA), but they are often limited when queries require cultural and visual information,…
Once Correct, Still Wrong: Counterfactual Hallucination in Multilingual Vision-Language Models
Basel Mousi, Fahim Dalvi, Shammur Chowdhury +2
Vision-language models (VLMs) can achieve high accuracy while still accepting culturally plausible but visually incorrect interpretations. Existing hallucination benchmarks rarely…
NativQA Framework: Enabling LLMs and VLMs with Native, Local, and Everyday Knowledge
Firoj Alam, Md Arid Hasan, Sahinur Rahman Laskar +3
The rapid progress of large language models (LLMs) raises concerns about cultural bias, fairness, and performance in diverse languages and underrepresented regions. Addressing thes…
Beyond LLM-as-a-Judge: Deterministic Metrics for Multilingual Generative Text Evaluation
Firoj Alam, Gagan Bhatia, Sahinur Rahman Laskar +1
While Large Language Models (LLMs) are increasingly adopted as automated judges for evaluating generated text, their outputs are often costly, and highly sensitive to prompt design…