152 citations · 184 across the 10 of their papers we have counts for
8 papers · 1 filter
HSVLT: Hierarchical Scale-Aware Vision-Language Transformer for Multi-Label Image Classification
Shuyi Ouyang, Hongyi Wang, Ziwei Niu +6
The task of multi-label image classification involves recognizing multiple objects within a single image. Considering both valuable semantic information contained in the labels and…
A Survey on Domain Generalization for Medical Image Analysis
Ziwei Niu, Shuyi Ouyang, Shiao Xie +2
Medical Image Analysis (MedIA) has emerged as a crucial tool in computer-aided diagnosis systems, particularly with the advancement of deep learning (DL) in recent years. However,…
Memory-Inspired Temporal Prompt Interaction for Text-Image Classification
Xinyao Yu, Hao Sun, Ziwei Niu +4
In recent years, large-scale pre-trained multimodal models (LMM) generally emerge to integrate the vision and language modalities, achieving considerable success in various natural…
M2ORT: Many-To-One Regression Transformer for Spatial Transcriptomics Prediction from Histopathology Images
Hongyi Wang, Xiuju Du, Jing Liu +3
The advancement of Spatial Transcriptomics (ST) has facilitated the spatially-aware profiling of gene expressions based on histopathology images. Although ST data offers valuable i…
HAP: Structure-Aware Masked Image Modeling for Human-Centric Perception
Junkun Yuan, Xinyu Zhang, Hao Zhou +12
Model pre-training is essential in human-centric perception. In this paper, we first introduce masked image modeling (MIM) as a pre-training approach for this task. Upon revisiting…
Tailored Multi-Organ Segmentation with Model Adaptation and Ensemble
Jiahua Dong, Guohua Cheng, Yue Zhang +5
Multi-organ segmentation, which identifies and separates different organs in medical images, is a fundamental task in medical image analysis. Recently, the immense success of deep…