1 paper
Mohammad Zbeeb, Hasan Abed Al Kader Hammoud, Sina Mukalled +5
We present AraLingBench: a fully human annotated benchmark for evaluating the Arabic linguistic competence of large language models (LLMs). The benchmark spans five core categories…