4 citations · 16 across the 8 of their papers we have counts for
8 papers · 1 filter
CartoonSing: Unifying Human and Nonhuman Timbres in Singing Generation
Jionghao Han, Jiatong Shi, Zhuoyan Tao +4
Singing voice synthesis (SVS) and singing voice conversion (SVC) have achieved remarkable progress in generating natural-sounding human singing. However, existing systems are restr…
TOMI: Transforming and Organizing Music Ideas for Multi-Track Compositions with Full-Song Structure
Qi He, Gus Xia, Ziyu Wang
Hierarchical planning is a powerful approach to model long sequences structurally. Aside from considering hierarchies in the temporal structure of music, this paper explores an eve…
ChatMusician: Understanding and Generating Music Intrinsically with LLM
Ruibin Yuan, Hanfeng Lin, Yi Wang +32
While Large Language Models (LLMs) demonstrate impressive capabilities in text generation, we find that their ability has yet to be generalized to music, humanity's creative langua…
Motif-Centric Representation Learning for Symbolic Music
Yuxuan Wu, Roger B. Dannenberg, Gus Xia
Music motif, as a conceptual building block of composition, is crucial for music structure analysis and automatic composition. While human listeners can identify motifs easily, exi…
Polyffusion: A Diffusion Model for Polyphonic Score Generation with Internal and External Controls
Lejun Min, Junyan Jiang, Gus Xia +1
We propose Polyffusion, a diffusion model that generates polyphonic music scores by regarding music as image-like piano roll representations. The model is capable of controllable m…
Q&A: Query-Based Representation Learning for Multi-Track Symbolic Music re-Arrangement
Jingwei Zhao, Gus Xia, Ye Wang
Music rearrangement is a common music practice of reconstructing and reconceptualizing a piece using new composition or instrumentation styles, which is also an important task of a…