59 citations · 72 across the 8 of their papers we have counts for
3 papers · 1 filter
AE-Flow: AutoEncoder Normalizing Flow
Jakub Mosiński, Piotr Biliński, Thomas Merritt +2
Recently normalizing flows have been gaining traction in text-to-speech (TTS) and voice conversion (VC) due to their state-of-the-art (SOTA) performance. Normalizing flows are unsu…
Creating New Voices using Normalizing Flows
Piotr Bilinski, Thomas Merritt, Abdelhamid Ezzerg +5
Creating realistic and natural-sounding synthetic speech remains a big challenge for voice identities unseen during training. As there is growing interest in synthesizing voices of…
SCRAPS: Speech Contrastive Representations of Acoustic and Phonetic Spaces
Ivan Vallés-Pérez, Grzegorz Beringer, Piotr Bilinski +2
Numerous examples in the literature proved that deep learning models have the ability to work well with multimodal data. Recently, CLIP has enabled deep learning systems to learn s…