Showing cs.SDShow all
2 papers · 1 filter
cs.SD2024
Investigating Disentanglement in a Phoneme-level Speech Codec for Prosody Modeling
Sotirios Karapiperis, Nikolaos Ellinas, Alexandra Vioni +4
Most of the prevalent approaches in speech prosody modeling rely on learning global style representations in a continuous latent space which encode and transfer the attributes of r…
cs.SD2024
Cross-lingual Text-To-Speech with Flow-based Voice Conversion for Improved Pronunciation
Nikolaos Ellinas, Georgios Vamvoukakis, Konstantinos Markopoulos +7
This paper presents a method for end-to-end cross-lingual text-to-speech (TTS) which aims to preserve the target language's pronunciation regardless of the original speaker's langu…