activity
20242026
collaborators

12 papers

cs.CL2026

Mawqif-XT: An Arabic Benchmark Dataset for Cross-Target Stance Detection

Rasha Albalawi, Nuha Albadi, Hamzah Luqman +4

Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This paper presents the Mawqif-XT,…

cs.RO2026

RAG-Based Auto-Configuration for Industrial Fieldbus Devices

Aadil Gani Ganie, Saad Ezzini, Naveed Farooz Marazi

Industrial device commissioning requires engineers to manually extract hundreds of protocol-specific parameters from heterogeneous PDF manuals and transcribe them into supervisory…

cs.SE2026

Empirical Study for Structured Output Control in LLMs for Software Engineering

Yewei Song, Prateek Rajput, Tiezhu Sun +3

LLM-generated outputs in software engineering rarely exist in isolation. They must plug into toolchains, APIs, and data pipelines that impose strict, often organization-specific st…

cs.AI2026

LeGo-Code: Can Modular Curriculum Learning Advance Complex Code Generation? Insights from Text-to-SQL

Salmane Chafik, Saad Ezzini, Ismail Berrada

Recently, code-oriented large language models (LLMs) have demonstrated strong capabilities in translating natural language into executable code. Text-to-SQL is a significant applic…

cs.CL2026

Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG

Omar El Bachyr, Yewei Song, Saad Ezzini +5

PDF files are primarily intended for human reading rather than automated processing. In addition, the heterogeneous content of PDFs, such as text, tables, and images, poses signifi…

cs.CL2025

AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP

Ahmed Hasanaath, Aisha Alansari, Ahmed Ashraf +3

Large language models (LLMs) have shown remarkable progress in reasoning abilities and general natural language processing (NLP) tasks, yet their performance on Arabic data, charac…