3 papers
cs.SD2026
Stemphonic: All-at-once Flexible Multi-stem Music Generation
Shih-Lun Wu, Ge Zhu, Juan-Pablo Caceres +2
Music stem generation, the task of producing musically-synchronized and isolated instrument audio clips, offers the potential of greater user control and better alignment with musi…
cs.SD2025
MIDI-LLM: Improving Text-to-MIDI Music Generation via Adapting Large Language Models
Shih-Lun Wu, Yoon Kim, Dave Carlton +3
We present MIDI-LLM, a recipe that improves multitrack text-to-MIDI generation via adapting Large Language Models (LLMs). MIDI-LLM expands an LLM's text vocabulary to include MIDI…
cs.SD2025
Streaming Generation for Music Accompaniment
Yusong Wu, Mason Wang, Heidi Lei +5
Music generation models can produce high-fidelity coherent accompaniment given complete audio input, but are limited to editing and loop-based workflows. We study real-time audio-t…