4 papers
CP-Env: Evaluating Large Language Models on Clinical Pathways in a Controllable Hospital Environment
Yakun Zhu, Zhongzhen Huang, Qianhan Feng +5
Medical care follows complex clinical pathways that extend beyond isolated physician-patient encounters, emphasizing decision-making and transitions between different stages. Curre…
UCAgents: Unidirectional Convergence for Visual Evidence Anchored Multi-Agent Medical Decision-Making
Qianhan Feng, Zhongzhen Huang, Yakun Zhu +2
Vision-Language Models (VLMs) show promise in medical diagnosis, yet suffer from reasoning detachment, where linguistically fluent explanations drift from verifiable image evidence…
SAP-Bench: Benchmarking Multimodal Large Language Models in Surgical Action Planning
Mengya Xu, Zhongzhen Huang, Dillan Imans +3
Effective evaluation is critical for driving advancements in MLLM research. The surgical action planning (SAP) task, which aims to generate future action sequences from visual inpu…
Surgical Action Planning with Large Language Models
Mengya Xu, Zhongzhen Huang, Jie Zhang +2
In robot-assisted minimally invasive surgery, we introduce the Surgical Action Planning (SAP) task, which generates future action plans from visual inputs to address the absence of…