6 papers · 1 filter
Paired Uterine Whole-Slide Images and Pathology Reports for Multimodal Computational Pathology
Han Li, Jingsong Liu, Ayako Ura +14
Uterine diseases represent an important category of gynecologic pathology and require accurate histopathological assessment for diagnosis and treatment planning. Whole-slide images…
QMix: Quality-aware Learning with Mixed Noise for Robust Retinal Disease Diagnosis
Junlin Hou, Jilan Xu, Rui Feng +1
Due to the complexity of medical image acquisition and the difficulty of annotation, medical image datasets inevitably contain noise. Noisy data with wrong labels affects the robus…
EgoExo-Gen: Ego-centric Video Prediction by Watching Exo-centric Videos
Jilan Xu, Yifei Huang, Baoqi Pei +6
Generating videos in the first-person perspective has broad application prospects in the field of augmented reality and embodied intelligence. In this work, we explore the cross-vi…
Concept-Attention Whitening for Interpretable Skin Lesion Diagnosis
Junlin Hou, Jilan Xu, Hao Chen
The black-box nature of deep learning models has raised concerns about their interpretability for successful deployment in real-world clinical applications. To address the concerns…
MOSMOS: Multi-organ segmentation facilitated by medical report supervision
Weiwei Tian, Xinyu Huang, Junlin Hou +6
Owing to a large amount of multi-modal data in modern medical systems, such as medical images and reports, Medical Vision-Language Pre-training (Med-VLP) has demonstrated incredibl…
Retrieval-Augmented Egocentric Video Captioning
Jilan Xu, Yifei Huang, Junlin Hou +4
Understanding human actions from videos of first-person view poses significant challenges. Most prior approaches explore representation learning on egocentric videos only, while ov…