1 paper
Hao Zhou, Xiaobao Guo, Yuzhe Zhu +1
Propelled by the breakthrough in deep generative models, audio-to-image generation has emerged as a pivotal cross-modal task that converts complex auditory signals into rich visual…