2 papers
cs.CV2025
DGTRSD & DGTRS-CLIP: A Dual-Granularity Remote Sensing Image-Text Dataset and Vision Language Foundation Model for Alignment
Weizhi Chen, Yupeng Deng, Jin Wei +7
Vision Language Foundation Models based on CLIP architecture for remote sensing primarily rely on short text captions, which often result in incomplete semantic representations. Al…
cs.CV2025
IRSAMap:Towards Large-Scale, High-Resolution Land Cover Map Vectorization
Yu Meng, Ligao Deng, Zhihao Xi +9
With the enhancement of remote sensing image resolution and the rapid advancement of deep learning, land cover mapping is transitioning from pixel-level segmentation to object-base…