Showing cs.SDShow all
3 papers · 1 filter
cs.SD2026
SAME: A Semantically-Aligned Music Autoencoder
Julian D. Parker, Zach Evans, CJ Carr +4
Latent representations are at the heart of the majority of modern generative models. In the audio domain they are typically produced by a neural-audio-codec autoencoder. In this wo…
cs.SD2026
Stable Audio 3
Zach Evans, Julian D. Parker, Matthew Rice +4
Stable Audio 3 is a family of fast latent diffusion models (small, medium, large) for variable-length audio generation and editing. Since our models can generate several minutes of…
cs.SD2023
General Purpose Audio Effect Removal
Matthew Rice, Christian J. Steinmetz, George Fazekas +1
Although the design and application of audio effects is well understood, the inverse problem of removing these effects is significantly more challenging and far less studied. Recen…