activity
20242026
collaborators

6 papers

cs.CV2026

SemiGDA: Generative Dual-distribution Alignment for Semi-Supervised Medical Image Segmentation

Kaiwen Huang, Yi Zhou, Yizhe Zhang +2

Semi-supervised learning addresses label scarcity and high annotation costs in medical image segmentation by exploiting the latent information in unlabeled data to enhance model pe…

cs.CV2025

AEM: Attention Entropy Maximization for Multiple Instance Learning based Whole Slide Image Classification

Yunlong Zhang, Honglin Li, Yunxuan Sun +4

Multiple Instance Learning (MIL) effectively analyzes whole slide images but faces overfitting due to attention over-concentration. While existing solutions rely on complex archite…

cs.CV2024

Large-scale cervical precancerous screening via AI-assisted cytology whole slide image analysis

Honglin Li, Yusuan Sun, Chenglu Zhu +8

Cervical Cancer continues to be the leading gynecological malignancy, posing a persistent threat to women's health on a global scale. Early screening via cytology Whole Slide Image…

cs.CV2024

PathGen-1.6M: 1.6 Million Pathology Image-text Pairs Generation through Multi-agent Collaboration

Yuxuan Sun, Yunlong Zhang, Yixuan Si +7

Vision Language Models (VLMs) like CLIP have attracted substantial attention in pathology, serving as backbones for applications such as zero-shot image classification and Whole Sl…

cs.CV2024

Benchmarking PathCLIP for Pathology Image Analysis

Sunyi Zheng, Xiaonan Cui, Yuxuan Sun +7

Accurate image classification and retrieval are of importance for clinical diagnosis and treatment decision-making. The recent contrastive language-image pretraining (CLIP) model h…

cs.CV2024

PathMMU: A Massive Multimodal Expert-Level Benchmark for Understanding and Reasoning in Pathology

Yuxuan Sun, Hao Wu, Chenglu Zhu +11

The emergence of large multimodal models has unlocked remarkable potential in AI, particularly in pathology. However, the lack of specialized, high-quality benchmark impeded their…