6 papers
SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation
Kaiwen Huang, Yi Zhou, Yizhe Zhang +2
Semi-supervised learning addresses label scarcity and high annotation costs in medical image segmentation by exploiting the latent information in unlabeled data to enhance model pe…
AEM: Attention Entropy Maximization for Multiple Instance Learning based Whole Slide Image Classification
Yunlong Zhang, Honglin Li, Yunxuan Sun +4
Multiple Instance Learning (MIL) effectively analyzes whole slide images but faces overfitting due to attention over-concentration. While existing solutions rely on complex archite…
Large-scale cervical precancerous screening via AI-assisted cytology whole slide image analysis
Honglin Li, Yusuan Sun, Chenglu Zhu +8
Cervical Cancer continues to be the leading gynecological malignancy, posing a persistent threat to women's health on a global scale. Early screening via cytology Whole Slide Image…
PathGen-1.6M: 1.6 Million Pathology Image-text Pairs Generation through Multi-agent Collaboration
Yuxuan Sun, Yunlong Zhang, Yixuan Si +7
Vision Language Models (VLMs) like CLIP have attracted substantial attention in pathology, serving as backbones for applications such as zero-shot image classification and Whole Sl…
Benchmarking PathCLIP for Pathology Image Analysis
Sunyi Zheng, Xiaonan Cui, Yuxuan Sun +7
Accurate image classification and retrieval are of importance for clinical diagnosis and treatment decision-making. The recent contrastive language-image pretraining (CLIP) model h…
PathMMU: A Massive Multimodal Expert-Level Benchmark for Understanding and Reasoning in Pathology
Yuxuan Sun, Hao Wu, Chenglu Zhu +11
The emergence of large multimodal models has unlocked remarkable potential in AI, particularly in pathology. However, the lack of specialized, high-quality benchmark impeded their…