2 papers
cs.CV2025
MEJO: MLLM-Engaged Surgical Triplet Recognition via Inter- and Intra-Task Joint Optimization
Yiyi Zhang, Yuchen Yuan, Ying Zheng +4
Surgical triplet recognition, which involves identifying instrument, verb, target, and their combinations, is a complex surgical scene understanding challenge plagued by long-taile…
eess.IV2025
Cardiac-CLIP: A Vision-Language Foundation Model for 3D Cardiac CT Images
Yutao Hu, Ying Zheng, Shumei Miao +20
Foundation models have demonstrated remarkable potential in medical domain. However, their application to complex cardiovascular diagnostics remains underexplored. In this paper, w…