4 papers
TimeSenCLIP: A Time Series Vision-Language Model for Remote Sensing
Pallavi Jain, Diego Marcos, Dino Ienco +2
Vision-language models (VLMs) have shown significant promise in remote sensing applications, particularly for land-use and land-cover (LULC) mapping via zero-shot classification an…
Atomizer: Generalizing to new modalities by breaking satellite images down to a set of scalars
Hugo Riffaud de Turckheim, Sylvain Lobry, Roberto Interdonato +1
The growing number of Earth observation satellites has led to increasingly diverse remote sensing data, with varying spatial, spectral, and temporal configurations. Most existing m…
Two-stage Vision Transformers and Hard Masking offer Robust Object Representations
Ananthu Aniraj, Cassio F. Dantas, Dino Ienco +1
Context can strongly affect object representations, sometimes leading to undesired biases, particularly when objects appear in out-of-distribution backgrounds at inference. At the…
SenCLIP: Enhancing zero-shot land-use mapping for Sentinel-2 with ground-level prompting
Pallavi Jain, Dino Ienco, Roberto Interdonato +2
Pre-trained vision-language models (VLMs), such as CLIP, demonstrate impressive zero-shot classification capabilities with free-form prompts and even show some generalization in sp…