activity
20232026
most citedFoundation Models for Music: A Survey

8 citations · 23 across the 24 of their papers we have counts for

collaborators
Showing cs.SDShow all

6 papers · 1 filter

cs.SD2024

SongTrans: An unified song transcription and alignment method for lyrics and notes

Siwei Wu, Jinzheng He, Ruibin Yuan +5

The quantity of processed data is crucial for advancing the field of singing voice synthesis. While there are tools available for lyric or note transcription tasks, they all need p…

cs.SD20248 cited

Foundation Models for Music: A Survey

Yinghao Ma, Anders Øland, Anton Ragni +39

In recent years, foundation models (FMs) such as large language models (LLMs) and latent diffusion models (LDMs) have profoundly impacted diverse sectors, including music. This com…

cs.SD20246 cited

ComposerX: Multi-Agent Symbolic Music Composition with LLMs

Qixin Deng, Qikai Yang, Ruibin Yuan +16

Music composition represents the creative side of humanity, and itself is a complex task that requires abilities to understand and generate information with long dependency and har…

cs.SD2024

MuPT: A Generative Symbolic Music Pretrained Transformer

Xingwei Qu, Yuelin Bai, Yinghao Ma +25

In this paper, we explore the application of Large Language Models (LLMs) to the pre-training of music. While the prevalent use of MIDI in music modeling is well-established, our f…

cs.SD20244 cited

ChatMusician: Understanding and Generating Music Intrinsically with LLM

Ruibin Yuan, Hanfeng Lin, Yi Wang +32

While Large Language Models (LLMs) demonstrate impressive capabilities in text generation, we find that their ability has yet to be generalized to music, humanity's creative langua…

cs.SD2023

Audio Contrastive-based Fine-tuning: Decoupling Representation Learning and Classification

Yang Wang, Qibin Liang, Chenghao Xiao +3

Standard fine-tuning of pre-trained audio models couples representation learning with classifier training, which can obscure the true quality of the learned representations. In thi…