3 papers
cs.CV2026
MIND: Multimodal Intent-Driven Network via Diffusion Transformers for Medical Image Fusion
Yunzhan Fu, Xiangyu Shen, Yifei Sun +3
Medical image fusion aims to integrate complementary information from diverse imaging modalities to support clinical diagnosis. Existing methods typically apply uniform fusion rule…
cs.CV2026
SCALPEL: Semantic Cross-modal Alignment via LLM-Powered Encoder Learning for Medical Vision-Language Representation
Yunzhan Fu, Enyu Bao, Xiangyu Shen +4
Vision-language pre-training (VLP) serves as a cornerstone for medical multimodal representation learning. However, existing medical VLP frameworks are often constrained by the lim…
eess.IV2025
ITCFN: Incomplete Triple-Modal Co-Attention Fusion Network for Mild Cognitive Impairment Conversion Prediction
Xiangyang Hu, Xiangyu Shen, Yifei Sun +8
Alzheimer's disease (AD) is a common neurodegenerative disease among the elderly. Early prediction and timely intervention of its prodromal stage, mild cognitive impairment (MCI),…