Showing cs.SDShow all
2 papers · 1 filter
cs.SD2024
M2M-Gen: A Multimodal Framework for Automated Background Music Generation in Japanese Manga Using Large Language Models
Megha Sharma, Muhammad Taimoor Haseeb, Gus Xia +1
This paper introduces M2M Gen, a multi modal framework for generating background music tailored to Japanese manga. The key challenges in this task are the lack of an available data…
cs.SD2024
Foundation Models for Music: A Survey
Yinghao Ma, Anders Ãland, Anton Ragni +39
In recent years, foundation models (FMs) such as large language models (LLMs) and latent diffusion models (LDMs) have profoundly impacted diverse sectors, including music. This com…