From the 1 of 16 linked papers with an AI index.
16 papers
RSVideo: Are Your Vision-Language Models Ready for Remote Sensing Videos?
Hongjie Zhou, Shiqin Wang, Haoyang Chen +5
Remote-sensing videos enable real-time observation of changes in target attributes, short-term activities, and scene evolution. They record motion, actions, interactions, and scene…
PNEC-Mamba: Prototype-Guided Positive-Negative Evidence Calibration for Hyperspectral Image Classification
Mingzhen Xu, Can Xu, Di Wang +2
In real-world hyperspectral scenes, pixel representations are often ambiguous due to factors such as spectral similarity, mixed pixels, and local context interference, which may si…
MBTI: A Multi-Branch Efficient Fine-Tuning Framework for Hyperspectral Image Classification with Foundation Models
Mingzhen Xu, Haonan Guo, Di Wang +9
The paper introduces MBTI, a multi-branch fine‑tuning framework that adapts hyperspectral foundation models to classification tasks while preserving full‑band spectral information…
Degradation-Aware Metric Prompting for Hyperspectral Image Restoration
Binfeng Wang, Di Wang, Haonan Guo +2
Unified hyperspectral image (HSI) restoration aims to recover diverse degradations within a single model. However, current methods often rely on impractical explicit priors or opaq…
Any2Any: Unified Arbitrary Modality Translation for Remote Sensing
Haoyang Chen, Jing Zhang, Hebaixu Wang +7
Multi-modal remote sensing imagery provides complementary observations of the same geographic scene, yet such observations are frequently incomplete in practice. Existing cross-mod…
VLRS-Bench: A Vision-Language Reasoning Benchmark for Remote Sensing
Zhiming Luo, Di Wang, Haonan Guo +2
Recent advancements in Multimodal Large Language Models (MLLMs) have enabled complex reasoning. However, existing remote sensing (RS) benchmarks remain heavily biased toward percep…