6 citations · 9 across the 11 of their papers we have counts for
7 papers · 1 filter
Gender Bias in Large Language Models for Healthcare: Assignment Consistency and Clinical Implications
Mingxuan Liu, Yuhe Ke, Wentao Zhu +9
The integration of large language models (LLMs) into healthcare holds promise to enhance clinical decision-making, yet their susceptibility to biases remains a critical concern. Ge…
MedBrowseComp: Benchmarking Medical Deep Research and Computer Use
Shan Chen, Pedro Moreira, Yuxin Xiao +6
Large language models (LLMs) are increasingly envisioned as decision-support tools in clinical practice, yet safe clinical reasoning demands integrating heterogeneous knowledge bas…
TheBlueScrubs-v1, a comprehensive curated medical dataset derived from the internet
Luis Felipe, Carlos Garcia, Issam El Naqa +6
The need for robust and diverse data sets to train clinical large language models (cLLMs) is critical given that currently available public repositories often prove too limited in…
Multi-OphthaLingua: A Multilingual Benchmark for Assessing and Debiasing LLM Ophthalmological QA in LMICs
David Restrepo, Chenwei Wu, Zhengxu Tang +14
Current ophthalmology clinical workflows are plagued by over-referrals, long waits, and complex and heterogeneous medical records. Large language models (LLMs) present a promising…
The use of large language models to enhance cancer clinical trial educational materials
Mingye Gao, Aman Varshney, Shan Chen +15
Cancer clinical trials often face challenges in recruitment and engagement due to a lack of participant-facing informational and educational resources. This study investigated the…
Language Models are Surprisingly Fragile to Drug Names in Biomedical Benchmarks
Jack Gallifant, Shan Chen, Pedro Moreira +7
Medical knowledge is context-dependent and requires consistent reasoning across various natural language expressions of semantically equivalent phrases. This is particularly crucia…