Showing cs.CVShow all
3 papers · 1 filter
cs.CV2026
Unlocking Zero-Shot Geospatial Reasoning via Indirect Rewards
Chenhui Xu, Fuxun Yu, Michael J. Bianco +15
Training robust reasoning vision-language models (VLMs) in rare domains (such as geospatial) is fundamentally constrained by supervision scarcity. While raw geospatial imagery is a…
cs.CV2025
Geospatial Foundational Embedder: Top-1 Winning Solution on EarthVision Embed2Scale Challenge (CVPR 2025)
Zirui Xu, Raphael Tang, Mike Bianco +4
EarthVision Embed2Scale challenge (CVPR 2025) aims to develop foundational geospatial models to embed SSL4EO-S12 hyperspectral geospatial data cubes into embedding vectors that fac…
cs.CV2024
Understanding Retrieval Robustness for Retrieval-Augmented Image Captioning
Wenyan Li, Jiaang Li, Rita Ramos +2
Recent advances in retrieval-augmented models for image captioning highlight the benefit of retrieving related captions for efficient, lightweight models with strong domain-transfe…