7 papers
P2E-VQ: ECG-linked representation augmentation for PPG via discrete patch retrieval
Zhongli Wu, Zhuangzhi Gao, He Zhao +8
Photoplethysmography (PPG) is widely used in consumer wearables because of its low cost and ease of acquisition. However, unlike electrocardiography (ECG), PPG measures peripheral…
HadBalance: A Plug-and-Play Unified Global Geometric Prior Framework for Generalizable Biomedical Segmentation
Zhuangzhi Gao, Feixiang Zhou, He Zhao +11
Precise biomedical image segmentation is crucial for clinical diagnosis. Geometric cues (e.g., boundary, shape, and topology) can improve structural consistency, yet most are task-…
Disentangled Fine-Grained Prototype Learning for Incomplete Image-Tabular Classification
Feixiang Zhou, Jianyang Xie, Zhuangzhi Gao +10
The missing-modality problem poses a significant challenge in image-tabular multimodal learning across a wide range of multimedia applications, including product understanding, rec…
Interpreting V1 Population Activity via Image-Neural Latent Representation Alignment
Xin Wang, Zhuangzhi Gao, Hongyi Qin +3
Understanding the neural mechanisms underlying visual computation has long been a central challenge in neuroscience. Recent alignment based approaches have improved the accuracy of…
Leveraging Persistence Image to Enhance Robustness and Performance in Curvilinear Structure Segmentation
Zhuangzhi Gao, Feixiang Zhou, He Zhao +8
Segmenting curvilinear structures in medical images is essential for analyzing morphological patterns in clinical applications. Integrating topological properties, such as connecti…
RetiBridge: Bridging Quantitative Retinal Biomarkers and Qualitative Diagnosis with a Knowledge-Guided Multimodal Large Language Model
Zhuangzhi Gao, Hongyi Qin, He Zhao +10
Retinal biomarkers captured by color fundus photography and optical coherence tomography provide clinically valuable evidence for both ocular and systemic diseases. Multimodal larg…