most citedViLa-MIL: Dual-scale Vision-Language Multiple Instance Learning for Whole Slide Image Classification

1 citations · 3 across the 6 of their papers we have counts for

collaborators

8 papers

cs.CV20251 cited

Prompting Lipschitz-constrained network for multiple-in-one sparse-view CT reconstruction

Baoshun Shi, Ke Jiang, Qiusheng Lian +2

Despite significant advancements in deep learning-based sparse-view computed tomography (SVCT) reconstruction algorithms, these methods still encounter two primary limitations: (i)…

cs.CV2025

GEMeX-RMCoT: An Enhanced Med-VQA Dataset for Region-Aware Multimodal Chain-of-Thought Reasoning

Bo Liu, Xiangyu Zhao, Along He +3

Medical visual question answering aims to support clinical decision-making by enabling models to answer natural language questions based on medical images. While recent advances in…

cs.CV2025

MExD: An Expert-Infused Diffusion Model for Whole-Slide Image Classification

Jianwei Zhao, Xin Li, Fan Yang +5

Whole Slide Image (WSI) classification poses unique challenges due to the vast image size and numerous non-informative regions, which introduce noise and cause data imbalance durin…

cs.CV2025

Hierarchical Context Transformer for Multi-level Semantic Scene Understanding

Luoying Hao, Yan Hu, Yang Yue +4

A comprehensive and explicit understanding of surgical scenes plays a vital role in developing context-aware computer-assisted systems in the operating theatre. However, few works…

cs.CV20251 cited

ViLa-MIL: Dual-scale Vision-Language Multiple Instance Learning for Whole Slide Image Classification

Jiangbo Shi, Chen Li, Tieliang Gong +2

Multiple instance learning (MIL)-based framework has become the mainstream for processing the whole slide image (WSI) with giga-pixel size and hierarchical image context in digital…

eess.IV20251 cited

Fundus Image Quality Assessment and Enhancement: a Systematic Review

Heng Li, Haojin Li, Mingyang Ou +5

As an affordable and convenient eye scan, fundus photography holds the potential for preventing vision impairment, especially in resource-limited regions. However, fundus image deg…