10 citations · 14 across the 2 of their papers we have counts for
3 papers
MedHELM: Holistic Evaluation of Large Language Models for Medical Tasks
Suhana Bedi, Hejie Cui, Miguel Fuentes +78
While large language models (LLMs) achieve near-perfect scores on medical licensing exams, these evaluations inadequately reflect the complexity and diversity of real-world clinica…
Standing on FURM ground -- A framework for evaluating Fair, Useful, and Reliable AI Models in healthcare systems
Alison Callahan, Duncan McElfresh, Juan M. Banda +21
The impact of using artificial intelligence (AI) to guide patient care or operational processes is an interplay of the AI model's output, the decision-making protocol based on that…
A Multi-Center Study on the Adaptability of a Shared Foundation Model for Electronic Health Records
Lin Lawrence Guo, Jason Fries, Ethan Steinberg +6
Foundation models hold promise for transforming AI in healthcare by providing modular components that are easily adaptable to downstream healthcare tasks, making AI development mor…