1 citations · 1 across the 2 of their papers we have counts for
2 papers
eess.AS2024
Scaling Transformers for Low-Bitrate High-Quality Speech Coding
Julian D Parker, Anton Smirnov, Jordi Pons +4
The tokenization of speech with neural audio codec models is a vital part of modern AI pipelines for the generation or understanding of speech, alone or in a multimodal context. Tr…
cs.SD2023★ 1 cited
GTR-CTRL: Instrument and Genre Conditioning for Guitar-Focused Music Generation with Transformers
Pedro Sarmento, Adarsh Kumar, Yu-Hua Chen +3
Recently, symbolic music generation with deep learning techniques has witnessed steady improvements. Most works on this topic focus on MIDI representations, but less attention has…