3 papers
cs.CV2026
Exploring the Potential of Contrastive Language-Image Pre-training for Multi-Source Remote Sensing Data
Xiangyang Miao, Kelu Yao, Yekai Huang +7
Contrastive language-image learning (CLIP) has become a key paradigm for remote sensing vision-language understanding. However, existing remote sensing contrastive learning methods…
cs.CV2026
Lightweight Adaptation of General-Purpose VLMs for Multispectral and SAR Image Understanding
Shanji Liu, Kelu Yao, Junxiao Xue +5
General-purpose vision-language models (VLMs) now support strong visual recognition, instruction following, and generation. However, most pretrained visual encoders are built aroun…
cs.CV2025
Forensics-Bench: A Comprehensive Forgery Detection Benchmark Suite for Large Vision Language Models
Jin Wang, Chenghui Lv, Xian Li +6
Recently, the rapid development of AIGC has significantly boosted the diversities of fake media spread in the Internet, posing unprecedented threats to social security, politics, l…