most citedSherkala-Chat: Building a State-of-the-Art LLM for Kazakh in a Moderately Resourced Setting

1 citations · 4 across the 38 of their papers we have counts for

collaborators

40 papers

cs.CL2026

Right Frame, Wrong Rule: Cultural Cues Expose the Financial Knowledge Gap They Were Meant to Close

Rania Elbadry, Ahmed Heakl, Saeed Almheiri +12

When a question has valid answers under different normative frameworks, a language model must decide which framework to use and whether it can answer correctly within it. We call t…

cs.CL2026

CultureTalk-ID: A Multi-Task Dialogue Benchmark for Cultural Commonsense in Indonesian Local Languages

Muhammad Dehan Al Kautsar, Salsabila Pranida, Bilal Elbouardi +1

Culture is lived through conversation, yet existing Indonesian cultural commonsense benchmarks evaluate LLMs on short and isolated prompts, stripping away the dialogic context in w…

cs.CL2026

Jais 2: A Family of Arabic-Centric Open Large Language Models

Mohamed Anwar, Abed Alhakim Freihat, George Ibrahim +57

Jais 2 is a family of Arabic-centric large language models developed jointly by MBZUAI, Cerebras, and Inception, designed to advance Arabic-centric language modeling, with strong p…

cs.CV2026

Sci-Rho: A Multilingual Visually-Grounded Symbolic Benchmark for STEM Problems

Muhammad Falensi Azmi, Ikhlasul Akmal Hanif, Vallerie Alexandra Putra +3

Symbolic benchmarks have emerged as a key approach to assess model robustness under minor modifications to STEM-related questions. However, existing symbolic benchmarks mostly rema…

cs.CL2026

Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resource Languages

Saeed Almheiri, Bilal Elbouardi, Salsabila Zahirah Pranida +16

Idiomatic expressions pose a major challenge for multilingual NLP because their meanings shift between figurative and literal usage, often requiring context for accurate interpreta…

cs.CL2026

IndoBias: A Dual Track Culturally Grounded Benchmark for LLMs Bias Evaluation in Indonesian Languages

Ikhlasul Akmal Hanif, Muhammad Falensi Azmi, Filbert Aurelian Tjiaranata +2

Despite being home to more than 1300 ethnic groups and 700 indigenous languages, bias in Large Language Models has not been fully studied in Indonesia, thus leaving a critical gap…