4 papers
Multimodal Interpretation of Remote Sensing Images: Dynamic Resolution Input Strategy and Multi-scale Vision-Language Alignment Mechanism
Siyu Zhang, Lianlei Shan, Runhe Qiu
Multimodal fusion of remote sensing images serves as a core technology for overcoming the limitations of single-source data and improving the accuracy of surface information extrac…
Building Lightweight Semantic Segmentation Models for Aerial Images Using Dual Relation Distillation
Minglong Li, Lianlei Shan, Weiqiang Wang +3
Recently, there have been significant improvements in the accuracy of CNN models for semantic segmentation. However, these models are often heavy and suffer from low inference spee…
Synthetic Lung X-ray Generation through Cross-Attention and Affinity Transformation
Ruochen Pi, Lianlei Shan
Collecting and annotating medical images is a time-consuming and resource-intensive task. However, generating synthetic data through models such as Diffusion offers a cost-effectiv…
Edge-guided and Class-balanced Active Learning for Semantic Segmentation of Aerial Images
Lianlei Shan, Weiqiang Wang, Ke Lv +1
Semantic segmentation requires pixel-level annotation, which is time-consuming. Active Learning (AL) is a promising method for reducing data annotation costs. Due to the gap betwee…