12 papers
Mawqif-XT: An Arabic Benchmark Dataset for Cross-Target Stance Detection
Rasha Albalawi, Nuha Albadi, Hamzah Luqman +4
Publicly available Arabic datasets for target-specific stance detection remain limited, particularly for evaluating cross-target generalization. This paper presents the Mawqif-XT,…
RAG-Based Auto-Configuration for Industrial Fieldbus Devices
Aadil Gani Ganie, Saad Ezzini, Naveed Farooz Marazi
Industrial device commissioning requires engineers to manually extract hundreds of protocol-specific parameters from heterogeneous PDF manuals and transcribe them into supervisory…
Empirical Study for Structured Output Control in LLMs for Software Engineering
Yewei Song, Prateek Rajput, Tiezhu Sun +3
LLM-generated outputs in software engineering rarely exist in isolation. They must plug into toolchains, APIs, and data pipelines that impose strict, often organization-specific st…
LeGo-Code: Can Modular Curriculum Learning Advance Complex Code Generation? Insights from Text-to-SQL
Salmane Chafik, Saad Ezzini, Ismail Berrada
Recently, code-oriented large language models (LLMs) have demonstrated strong capabilities in translating natural language into executable code. Text-to-SQL is a significant applic…
Empirical Evaluation of PDF Parsing and Chunking for Financial Question Answering with RAG
Omar El Bachyr, Yewei Song, Saad Ezzini +5
PDF files are primarily intended for human reading rather than automated processing. In addition, the heterogeneous content of PDFs, such as text, tables, and images, poses signifi…
AraReasoner: Evaluating Reasoning-Based LLMs for Arabic NLP
Ahmed Hasanaath, Aisha Alansari, Ahmed Ashraf +3
Large language models (LLMs) have shown remarkable progress in reasoning abilities and general natural language processing (NLP) tasks, yet their performance on Arabic data, charac…