2 papers
cs.CV2026
Beyond GSD-as-Token: Continuous Scale Conditioning for Remote Sensing VLMs
Song Zhang, Yanlong Chen, Yilin Li +4
Remote sensing vision-language models (RS-VLMs) face a fundamental mismatch with natural-image counterparts: the same geographic object exhibits radically different visual evidence…
cs.CV2025
Region-to-Region: Enhancing Generative Image Harmonization with Adaptive Regional Injection
Zhiqiu Zhang, Dongqi Fan, Mingjie Wang +3
The goal of image harmonization is to adjust the foreground in a composite image to achieve visual consistency with the background. Recently, latent diffusion model (LDM) are appli…