Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Beyond Pixels: A Training-Free, Text-to-Text Framework for Remote Sensing Image Retrieval
J. Xiao, Y. Guo, X. Zi +3
Semantic retrieval of remote sensing (RS) images is a critical task fundamentally challenged by the \textquote{semantic gap}, the discrepancy between a model's low-level visual fea…
cs.CV2025
RSVLM-QA: A Benchmark Dataset for Remote Sensing Vision Language Model-based Question Answering
Xing Zi, Jinghao Xiao, Yunxiao Shi +4
Visual Question Answering (VQA) in remote sensing (RS) is pivotal for interpreting Earth observation data. However, existing RS VQA datasets are constrained by limitations in annot…
cs.CV2025
AeroLite: Tag-Guided Lightweight Generation of Aerial Image Captions
Xing Zi, Tengjun Ni, Xianjing Fan +4
Accurate and automated captioning of aerial imagery is crucial for applications like environmental monitoring, urban planning, and disaster management. However, this task remains c…