2 citations · 3 across the 3 of their papers we have counts for
3 papers
eess.AS2023★ 2 cited
A Neural TTS System with Parallel Prosody Transfer from Unseen Speakers
Slava Shechtman, Raul Fernandez
Modern neural TTS systems are capable of generating natural and expressive speech when provided with sufficient amounts of training data. Such systems can be equipped with prosody-…
eess.AS2023★ 1 cited
Speak While You Think: Streaming Speech Synthesis During Text Generation
Avihu Dekel, Slava Shechtman, Raul Fernandez +3
Large Language Models (LLMs) demonstrate impressive capabilities, yet interaction with these models is mostly facilitated through text. Using Text-To-Speech to synthesize LLM outpu…
eess.AS2022
Transplantation of Conversational Speaking Style with Interjections in Sequence-to-Sequence Speech Synthesis
Raul Fernandez, David Haws, Guy Lorberbom +2
Sequence-to-Sequence Text-to-Speech architectures that directly generate low level acoustic features from phonetic sequences are known to produce natural and expressive speech when…