3 papers
cs.CL2026
When Large Language Models Fail in Healthcare: Evaluating Sensitivity to Prompt Variations
Mahdi Alkaeed
Large Language Models (LLMs) are increasingly used in healthcare for tasks such as clinical question answering, diagnosis support, and report summarization. Despite their promise,…
cs.CL2026
Same Patient, Different Words, Different Diagnosis? Evaluating Semantic Stability in Clinical LLMs
Mahdi Alkaeed, Adnan Qayyum, Nabeel Abo Kashreef +2
Large Language Models (LLMs) are increasingly used in clinical applications. However, their behavior remains highly sensitive to subtle linguistic variations, such as rephrasing or…
cs.CL2025
Open Foundation Models in Healthcare: Challenges, Paradoxes, and Opportunities with GenAI Driven Personalized Prescription
Mahdi Alkaeed, Sofiat Abioye, Adnan Qayyum +6
In response to the success of proprietary Large Language Models (LLMs) such as OpenAI's GPT-4, there is a growing interest in developing open, non-proprietary LLMs and AI foundatio…