1 paper
Tornike Karchkhadze, Mohammad Rasool Izadi, Ke Chen +2
Diffusion models have shown promising results in cross-modal generation tasks involving audio and music, such as text-to-sound and text-to-music generation. These text-controlled m…