13 citations
3 papers
cs.CV2025★ 2 cited
Efficient Medical Vision-Language Alignment Through Adapting Masked Vision Models
Chenyu Lian, Hong-Yu Zhou, Dongyun Liang +2
Medical vision-language alignment through cross-modal contrastive learning shows promising performance in image-text matching tasks, such as retrieval and zero-shot classification.…
eess.IV2021
Real-time landmark detection for precise endoscopic submucosal dissection via shape-aware relation network
Jiacheng Wang, Yueming Jin, Shuntian Cai +4
We propose a novel shape-aware relation network for accurate and real-time landmark detection in endoscopic submucosal dissection (ESD) surgery. This task is of great clinical sign…
cs.CV2021★ 13 cited
Efficient Global-Local Memory for Real-time Instrument Segmentation of Robotic Surgical Video
Jiacheng Wang, Yueming Jin, Liansheng Wang +3
Performing a real-time and accurate instrument segmentation from videos is of great significance for improving the performance of robotic-assisted surgery. We identify two importan…