3 papers
cs.CV2026
Continual Vision-Language Learning for Remote Sensing: Benchmarking and Analysis
Xingxing Weng, Ruifeng Ni, Chao Pang +5
Current remote sensing vision-language models (RS VLMs) demonstrate impressive performance in image interpretation but rely on static training data, limiting their ability to accom…
cs.CV2025
CADFormer: Fine-Grained Cross-modal Alignment and Decoding Transformer for Referring Remote Sensing Image Segmentation
Maofu Liu, Xin Jiang, Xiaokang Zhang
Referring Remote Sensing Image Segmentation (RRSIS) is a challenging task, aiming to segment specific target objects in remote sensing (RS) images based on a given language express…
cs.CV2025
Semantic-Spatial Feature Fusion with Dynamic Graph Refinement for Remote Sensing Image Captioning
Maofu Liu, Jiahui Liu, Xiaokang Zhang
Remote sensing image captioning aims to generate semantically accurate descriptions that are closely linked to the visual features of remote sensing images. Existing approaches typ…