5 papers
GeoViSTA: Geospatial Vision-Tabular Transformer for Multimodal Environment Representation
Yuhao Liu, Sadeer Al-Kindi, Ashok Veeraraghavan +1
Large-scale pretraining on Earth observation imagery has yielded powerful representations of the natural and built environment. However, most existing geospatial foundation models…
CoLVR: Enhancing Exploratory Latent Visual Reasoning via Contrastive Optimization
Ziyang Ding, Linjian Meng, Yiming Wu +3
Due to the potential for exploratory reasoning of Latent Visual Reasoning, recent works tend to enable MLLMs (Multimodal Large Language Models) to perform visual reasoning by propa…
Fast Amortized Fitting of Scientific Signals Across Time and Ensembles via Transferable Neural Fields
Sophia Zorek, Kushal Vyas, Yuhao Liu +3
Neural fields, also known as implicit neural representations (INRs), offer a powerful framework for modeling continuous geometry, but their effectiveness in high-dimensional scient…
Downscaling Extreme Precipitation with Wasserstein Regularized Diffusion
Yuhao Liu, James Doss-Gollin, Qiushi Dai +2
Understanding the risks posed by extreme rainfall events requires analysis of precipitation fields with high resolution (to assess localized hazards) and extensive historical cover…
Post-Hurricane Debris Segmentation Using Fine-Tuned Foundational Vision Models
Kooshan Amini, Yuhao Liu, Jamie Ellen Padgett +2
Timely and accurate detection of hurricane debris is critical for effective disaster response and community resilience. While post-disaster aerial imagery is readily available, rob…