4 papers
TAVR-VLM: Risk-Conditioned Causal Grounding for Hallucination-Resistant Report Generation
Zhixiang Lu, Xiwei Liu, Sifan Song +5
Transcatheter Aortic Valve Replacement (TAVR) planning requires meticulous multimodal reasoning. However, adapting Multimodal Large Language Models (MLLMs) to this high-stakes doma…
A novel Fourier Adjacency Transformer for advanced EEG emotion recognition
Jinfeng Wang, Yanhao Huang, Sifan Song +3
EEG emotion recognition faces significant hurdles due to noise interference, signal nonstationarity, and the inherent complexity of brain activity which make accurately emotion cla…
Contrastive Learning Via Equivariant Representation
Sifan Song, Jinfeng Wang, Qiaochu Zhao +6
Invariant Contrastive Learning (ICL) methods have achieved impressive performance across various domains. However, the absence of latent space representation for distortion (augmen…
DualStreamFoveaNet: A Dual Stream Fusion Architecture with Anatomical Awareness for Robust Fovea Localization
Sifan Song, Jinfeng Wang, Zilong Wang +4
Accurate fovea localization is essential for analyzing retinal diseases to prevent irreversible vision loss. While current deep learning-based methods outperform traditional ones,…