3 papers
cs.SD2026
MuScriptor: An Open Model for Multi-Instrument Music Transcription
Simon Rouard, Michael Krause, Axel Roebel +2
Existing methods for automatic music transcription are often limited to single-instrument recordings or fail on complex, real music mixes. Although previous work utilizes synthetic…
cs.SD2026
Continuous Audio Language Models
Simon Rouard, Manu Orsini, Axel Roebel +2
Audio Language Models (ALM) have emerged as the dominant paradigm for speech and music generation by representing audio as sequences of discrete tokens. Yet, unlike text tokens, wh…
cs.SD2025
MusicGen-Stem: Multi-stem music generation and edition through autoregressive modeling
Simon Rouard, Robin San Roman, Yossi Adi +1
While most music generation models generate a mixture of stems (in mono or stereo), we propose to train a multi-stem generative model with 3 stems (bass, drums and other) that lear…