6 papers
MedExpMem: Adapting Experience Memory for Differential Diagnosis
Qianhan Feng, Zhongzhen Huang, Yakun Zhu +4
Experienced physicians develop diagnostic expertise through clinical practice, acquiring not only disease knowledge but also the ability to differentiate confusable conditions. Cur…
CP-Env: Evaluating Large Language Models on Clinical Pathways in a Controllable Hospital Environment
Yakun Zhu, Zhongzhen Huang, Qianhan Feng +5
Medical care follows complex clinical pathways that extend beyond isolated physician-patient encounters, emphasizing decision-making and transitions between different stages. Curre…
UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-Making
Qianhan Feng, Zhongzhen Huang, Yakun Zhu +2
Vision-Language Models (VLMs) show promise in medical diagnosis, yet suffer from reasoning detachment, where linguistically fluent explanations drift from verifiable image evidence…
SAP-Bench: Benchmarking Multimodal Large Language Models in Surgical Action Planning
Mengya Xu, Zhongzhen Huang, Dillan Imans +3
Effective evaluation is critical for driving advancements in MLLM research. The surgical action planning (SAP) task, which aims to generate future action sequences from visual inpu…
Surgical Action Planning with Large Language Models
Mengya Xu, Zhongzhen Huang, Jie Zhang +2
In robot-assisted minimally invasive surgery, we introduce the Surgical Action Planning (SAP) task, which generates future action plans from visual inputs to address the absence of…
Grounded Knowledge-Enhanced Medical Vision-Language Pre-training for Chest X-Ray
Qiao Deng, Zhongzhen Huang, Yunqi Wang +6
Medical foundation models have the potential to revolutionize healthcare by providing robust and generalized representations of medical data. Medical vision-language pre-training h…