27 citations · 38 across the 4 of their papers we have counts for
4 papers
Ensemble prosody prediction for expressive speech synthesis
Tian Huey Teh, Vivian Hu, Devang S Ram Mohan +7
Generating expressive speech with rich and varied prosody continues to be a challenge for Text-to-Speech. Most efforts have focused on sophisticated neural architectures intended t…
ADEPT: A Dataset for Evaluating Prosody Transfer
Alexandra Torresquintero, Tian Huey Teh, Christopher G. R. Wallis +6
Text-to-speech is now able to achieve near-human naturalness and research focus has shifted to increasing expressivity. One popular method is to transfer the prosody from a referen…
Incremental Text to Speech for Neural Sequence-to-Sequence Models using Reinforcement Learning
Devang S Ram Mohan, Raphael Lenain, Lorenzo Foglianti +4
Modern approaches to text to speech require the entire input character sequence to be processed before any audio is synthesised. This latency limits the suitability of such models…
Phonological Features for 0-shot Multilingual Speech Synthesis
Marlene Staib, Tian Huey Teh, Alexandra Torresquintero +4
Code-switching---the intra-utterance use of multiple languages---is prevalent across the world. Within text-to-speech (TTS), multilingual models have been found to enable code-swit…