599 citations · 753 across the 4 of their papers we have counts for
Showing 2023Show all
2 papers · 1 filter
cs.SD2023★ 7 cited
VampNet: Music Generation via Masked Acoustic Token Modeling
Hugo Flores Garcia, Prem Seetharaman, Rithesh Kumar +1
We introduce VampNet, a masked acoustic token modeling approach to music synthesis, compression, inpainting, and variation. We use a variable masking schedule during training which…
cs.SD2023
High-Fidelity Audio Compression with Improved RVQGAN
Rithesh Kumar, Prem Seetharaman, Alejandro Luebs +2
Language models have been successfully used to model natural signals, such as images, speech, and music. A key component of these models is a high quality neural compression model…