activity
20242026
collaborators
Showing cs.CLShow all

5 papers · 1 filter

cs.CL2026

PersianMedQA: Evaluating Large Language Models on a Persian-English Bilingual Medical Question Answering Benchmark

Mohammad Javad Ranjbar Kalahroodi, Amirhossein Sheikholselami, Sepehr Karimi +3

Large Language Models (LLMs) have achieved remarkable performance on a wide range of Natural Language Processing (NLP) benchmarks, often surpassing human-level accuracy. However, t…

cs.CL2026

PARSA-Bench: A Comprehensive Persian Audio-Language Model Benchmark

Mohammad Javad Ranjbar Kalahroodi, Mohammad Amini, Parmis Bathayan +2

Persian poses unique audio understanding challenges through its classical poetry, traditional music, and pervasive code-switching - none captured by existing benchmarks. We introdu…

cs.CL2026

PersianPunc: A Large-Scale Dataset and BERT-Based Approach for Persian Punctuation Restoration

Mohammad Javad Ranjbar Kalahroodi, Heshaam Faili, Azadeh Shakery

Punctuation restoration is essential for improving the readability and downstream utility of automatic speech recognition (ASR) outputs, yet remains underexplored for Persian despi…

cs.CL2025

PolitiSky24: U.S. Political Bluesky Dataset with User Stance Labels

Peyman Rostami, Vahid Rahimzadeh, Ali Adibi +1

Stance detection identifies the viewpoint expressed in text toward a specific target, such as a political figure. While previous datasets have focused primarily on tweet-level stan…

cs.CL2025

PerCul: A Story-Driven Cultural Evaluation of LLMs in Persian

Erfan Moosavi Monazzah, Vahid Rahimzadeh, Yadollah Yaghoobzadeh +2

Large language models predominantly reflect Western cultures, largely due to the dominance of English-centric training data. This imbalance presents a significant challenge, as LLM…