2 papers
cs.SD2026
SounDiT: Geo-Contextual Soundscape-to-Landscape Generation
Junbo Wang, Haofeng Tan, Bowen Liao +7
Recent audio-to-image models have shown impressive performance in generating images of specific objects conditioned on their corresponding sounds. However, these models fail to rec…
cs.LG2026
Decorrelating the Future: Joint Frequency Domain Learning for Spatio-temporal Forecasting
Zepu Wang, Bowen Liao, Jeff +1
Standard direct forecasting models typically rely on point-wise objectives such as Mean Squared Error, which fail to capture the complex spatio-temporal dependencies inherent in gr…