2 papers
cs.CL2020
A comparison of self-supervised speech representations as input features for unsupervised acoustic word embeddings
Lisa van Staden, Herman Kamper
Many speech processing tasks involve measuring the acoustic similarity between speech segments. Acoustic word embeddings (AWE) allow for efficient comparisons by mapping speech seg…
cs.CL2019
Unsupervised acoustic unit discovery for speech synthesis using discrete latent-variable neural networks
Ryan Eloff, André Nortje, Benjamin van Niekerk +7
For our submission to the ZeroSpeech 2019 challenge, we apply discrete latent-variable neural networks to unlabelled speech and use the discovered units for speech synthesis. Unsup…