2 papers
cs.CL2025
Performance of Large Language Models in Answering Critical Care Medicine Questions
Mahmoud Alwakeel, Aditya Nagori, An-Kwok Ian Wong +3
Large Language Models have been tested on medical student-level questions, but their performance in specialized fields like Critical Care Medicine (CCM) is less explored. This stud…
cs.CL2025
Evaluating LLMs in Medicine: A Call for Rigor, Transparency
Mahmoud Alwakeel, Aditya Nagori, Vijay Krishnamoorthy +1
Objectives: To evaluate the current limitations of large language models (LLMs) in medical question answering, focusing on the quality of datasets used for their evaluation. Materi…