1 paper
Shansong Liu, Atin Sakkeer Hussain, Qilong Wu +2
Research on large language models has advanced significantly across text, speech, images, and videos. However, multi-modal music understanding and generation remain underexplored d…