3 papers
cs.SD2026
Latent Fourier Transform
Mason Wang, Cheng-Zhi Anna Huang
We introduce the Latent Fourier Transform (LatentFT), a framework that provides novel frequency-domain controls for generative music models. LatentFT combines a diffusion autoencod…
cs.SD2025
Streaming Generation for Music Accompaniment
Yusong Wu, Mason Wang, Heidi Lei +5
Music generation models can produce high-fidelity coherent accompaniment given complete audio input, but are limited to editing and loop-based workflows. We study real-time audio-t…
cs.SD2025
Subtractive Training for Music Stem Insertion using Latent Diffusion Models
Ivan Villa-Renteria, Mason L. Wang, Zachary Shah +4
We present Subtractive Training, a simple and novel method for synthesizing individual musical instrument stems given other instruments as context. This method pairs a dataset of c…