7 citations · 10 across the 3 of their papers we have counts for
3 papers · 1 filter
Improved parallel WaveGAN vocoder with perceptually weighted spectrogram loss
Eunwoo Song, Ryuichi Yamamoto, Min-Jae Hwang +3
This paper proposes a spectral-domain perceptual weighting technique for Parallel WaveGAN-based text-to-speech (TTS) systems. The recently proposed Parallel WaveGAN vocoder success…
Neural text-to-speech with a modeling-by-generation excitation vocoder
Eunwoo Song, Min-Jae Hwang, Ryuichi Yamamoto +3
This paper proposes a modeling-by-generation (MbG) excitation vocoder for a neural text-to-speech (TTS) system. Recently proposed neural excitation vocoders can realize qualified w…
Effective parameter estimation methods for an ExcitNet model in generative text-to-speech systems
Ohsung Kwon, Eunwoo Song, Jae-Min Kim +1
In this paper, we propose a high-quality generative text-to-speech (TTS) system using an effective spectrum and excitation estimation method. Our previous research verified the eff…