Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
CA-GCL: Cross-Anatomy Global-Local Contrastive Learning for Robust 3D Medical Image Understanding
Hanwen Zhang, Yao Liu, Die Dai +4
Fine-grained Vision-Language Pre-training (FVLP) demonstrates significant potential in 3D medical image understanding by aligning anatomy-level visual representations with correspo…
cs.CV2025
LarvSeg: Exploring Image Classification Data For Large Vocabulary Semantic Segmentation via Category-wise Attentive Classifier
Haojun Yu, Di Dai, Ziwei Zhao +3
Scaling up the vocabulary of semantic segmentation models is extremely challenging because annotating large-scale mask labels is labour-intensive and time-consuming. Recently, lang…
cs.CV2023
PRED: Pre-training via Semantic Rendering on LiDAR Point Clouds
Hao Yang, Haiyang Wang, Di Dai +1
Pre-training is crucial in 3D-related fields such as autonomous driving where point cloud annotation is costly and challenging. Many recent studies on point cloud pre-training, how…