3 citations · 3 across the 4 of their papers we have counts for
4 papers · 1 filter
ChatEarthNet: A Global-Scale Image-Text Dataset Empowering Vision-Language Geo-Foundation Models
Zhenghang Yuan, Zhitong Xiong, Lichao Mou +1
An in-depth comprehension of global land cover is essential in Earth observation, forming the foundation for a multitude of applications. Although remote sensing technology has adv…
Overcoming Language Bias in Remote Sensing Visual Question Answering via Adversarial Training
Zhenghang Yuan, Lichao Mou, Xiao Xiang Zhu
The Visual Question Answering (VQA) system offers a user-friendly interface and enables human-computer interaction. However, VQA models commonly face the challenge of language bias…
GAMUS: A Geometry-aware Multi-modal Semantic Segmentation Benchmark for Remote Sensing Data
Zhitong Xiong, Sining Chen, Yi Wang +2
Geometric information in the normalized digital surface models (nDSM) is highly correlated with the semantic class of the land cover. Exploiting two modalities (RGB and nDSM (heigh…
Multilingual Augmentation for Robust Visual Question Answering in Remote Sensing Images
Zhenghang Yuan, Lichao Mou, Xiao Xiang Zhu
Aiming at answering questions based on the content of remotely sensed images, visual question answering for remote sensing data (RSVQA) has attracted much attention nowadays. Howev…