3 papers
cs.LG2026
Mechanisms of Misgeneralization in Physical Sequence Modeling
Kento Nishi, Raphael Tang, Karun Kumar +2
Generative sequence models are often trained to plan motion in physical domains, from robotics to mechanical simulations. When constructing a dataset to train such a model, enginee…
cs.CV2026
Unlocking Zero-Shot Geospatial Reasoning via Indirect Rewards
Chenhui Xu, Fuxun Yu, Michael J. Bianco +15
Training robust reasoning vision-language models (VLMs) in rare domains (such as geospatial) is fundamentally constrained by supervision scarcity. While raw geospatial imagery is a…
cs.CV2025
Geospatial Foundational Embedder: Top-1 Winning Solution on EarthVision Embed2Scale Challenge (CVPR 2025)
Zirui Xu, Raphael Tang, Mike Bianco +4
EarthVision Embed2Scale challenge (CVPR 2025) aims to develop foundational geospatial models to embed SSL4EO-S12 hyperspectral geospatial data cubes into embedding vectors that fac…