1 paper · 1 filter
Junbo Wang, Haofeng Tan, Bowen Liao +7
Recent audio-to-image models have shown impressive performance in generating images of specific objects conditioned on their corresponding sounds. However, these models fail to rec…