4 papers · 1 filter
SalamahBench: Toward Standardized Safety Evaluation for Arabic Language Models
Omar Abdelnasser, Fatemah Alharbi, Khaled Khasawneh +2
Safety alignment in Language Models (LMs) is fundamental for trustworthy AI. However, while different stakeholders are trying to leverage Arabic Language Models (ALMs), systematic…
PsychiatryBench: A Multi-Task Benchmark for LLMs in Psychiatry
Aya E. Fouda, Abdelrahamn A. Hassan, Radwa J. Hanafy +1
Large language models (LLMs) offer significant potential in enhancing psychiatric practice, from improving diagnostic accuracy to streamlining clinical documentation and therapeuti…
Leveraging Audio and Text Modalities in Mental Health: A Study of LLMs Performance
Abdelrahman A. Ali, Aya E. Fouda, Radwa J. Hanafy +1
Mental health disorders are increasingly prevalent worldwide, creating an urgent need for innovative tools to support early diagnosis and intervention. This study explores the pote…
A Comprehensive Evaluation of Large Language Models on Mental Illnesses in Arabic Context
Noureldin Zahran, Aya E. Fouda, Radwa J. Hanafy +1
Mental health disorders pose a growing public health concern in the Arab world, emphasizing the need for accessible diagnostic and intervention tools. Large language models (LLMs)…