1 citations · 2 across the 3 of their papers we have counts for
2 papers
eess.AS2024
Scaling Transformers for Low-Bitrate High-Quality Speech Coding
Julian D Parker, Anton Smirnov, Jordi Pons +4
The tokenization of speech with neural audio codec models is a vital part of modern AI pipelines for the generation or understanding of speech, alone or in a multimodal context. Tr…
cs.SD2024★ 1 cited
Stable Audio Open
Zach Evans, Julian D. Parker, CJ Carr +3
Open generative models are vitally important for the community, allowing for fine-tunes and serving as baselines when presenting new models. However, most current text-to-audio mod…