3 papers
cs.CV2024
ChatEarthNet: A Global-Scale Image-Text Dataset Empowering Vision-Language Geo-Foundation Models
Zhenghang Yuan, Zhitong Xiong, Lichao Mou +1
An in-depth comprehension of global land cover is essential in Earth observation, forming the foundation for a multitude of applications. Although remote sensing technology has adv…
cs.CV2023
Overcoming Language Bias in Remote Sensing Visual Question Answering via Adversarial Training
Zhenghang Yuan, Lichao Mou, Xiao Xiang Zhu
The Visual Question Answering (VQA) system offers a user-friendly interface and enables human-computer interaction. However, VQA models commonly face the challenge of language bias…
cs.CV2023
Multilingual Augmentation for Robust Visual Question Answering in Remote Sensing Images
Zhenghang Yuan, Lichao Mou, Xiao Xiang Zhu
Aiming at answering questions based on the content of remotely sensed images, visual question answering for remote sensing data (RSVQA) has attracted much attention nowadays. Howev…