4 papers
Composer Vector: Style-steering Symbolic Music Generation in a Latent Space
Xunyi Jiang, Mingyang Yao, Jingyue Huang +1
Symbolic music generation has made significant progress, yet achieving fine-grained and flexible control over composer style remains challenging. Existing training-based methods fo…
Low-Resource Guidance for Controllable Latent Audio Diffusion
Zachary Novack, Zack Zukowski, CJ Carr +6
Generative audio requires fine-grained controllable outputs, yet most existing methods require model retraining on specific controls or inference-time controls (\textit{e.g.}, guid…
StylePitcher: Generating Style-Following and Expressive Pitch Curves for Versatile Singing Tasks
Jingyue Huang, Qihui Yang, Fei Yueh Chen +4
Existing pitch curve generators face two main challenges: they often neglect singer-specific expressiveness, reducing their ability to capture individual singing styles. And they a…
Are you really listening? Boosting Perceptual Awareness in Music-QA Benchmarks
Yongyi Zang, Sean O'Brien, Taylor Berg-Kirkpatrick +2
Large Audio Language Models (LALMs), where pretrained text LLMs are finetuned with audio input, have made remarkable progress in music understanding. However, current evaluation me…