3 papers
cs.CV2026
SCALPEL: Semantic Cross-modal Alignment via LLM-Powered Encoder Learning for Medical Vision-Language Representation
Yunzhan Fu, Enyu Bao, Xiangyu Shen +4
Vision-language pre-training (VLP) serves as a cornerstone for medical multimodal representation learning. However, existing medical VLP frameworks are often constrained by the lim…
cs.CV2026
GLeVE: Graph-Guided Lesion Grounding with Proposal Verification in 3D CT
Shuo Jiang, Yuhao Hong, Chunbo Jiang +9
Grounding radiology report descriptions to 3D CT volumes is essential for verifiable clinical interpretation, yet remains challenging due to the semantic-spatial gap between free-t…
cs.CV2026
R2AoP: Reliable and Robust Angle of Progression Estimation from Intrapartum Ultrasound
Yuanhan Wang, Yifei Chen, Beining Wu +7
Accurate estimation of the Angle of Progression (AoP) from intrapartum transperineal ultrasound is critical for objective assessment of labor progression, yet remains highly sensit…