6 papers
EHRBench: An Automated and Reliable EHR-based Benchmark for Clinical Decision Making with LLMs
Yuzhang Xie, Keqi Han, Yunpeng Xiao +7
Clinical decision-making (CDM) is central to real-world clinical workflows, where clinicians infer diagnoses, select treatments, or anticipate future health outcomes under incomple…
EpiQAL: Benchmarking Large Language Models in Epidemiological Question Answering and Reasoning
Mingyang Wei, Dehai Min, Zewen Liu +8
Reliable epidemiological reasoning requires synthesizing study evidence to infer disease burden, transmission dynamics, and intervention effects at the population level. Existing m…
MIRAGE: Knowledge Graph-Guided Cross-Cohort MRI Synthesis for Alzheimer's Disease Prediction
Guanchen Wu, Zhe Huang, Yuzhang Xie +6
Reliable Alzheimer's disease (AD) diagnosis increasingly relies on multimodal assessments combining structural Magnetic Resonance Imaging (MRI) and Electronic Health Records (EHR).…
Utilizing Large Language Models for Zero-Shot Medical Ontology Extension from Clinical Notes
Guanchen Wu, Yuzhang Xie, Huanwei Wu +4
Integrating novel medical concepts and relationships into existing ontologies can significantly enhance their coverage and utility for both biomedical research and clinical applica…
Towards Automatic Evaluation and Selection of PHI De-identification Models via Multi-Agent Collaboration
Guanchen Wu, Zuhui Chen, Yuzhang Xie +1
Protected health information (PHI) de-identification is critical for enabling the safe reuse of clinical notes, yet evaluating and comparing PHI de-identification models typically…
Large Language Model Empowered Privacy-Protected Framework for PHI Annotation in Clinical Notes
Guanchen Wu, Linzhi Zheng, Han Xie +7
The de-identification of private information in medical data is a crucial process to mitigate the risk of confidentiality breaches, particularly when patient personal details are n…