7 citations · 10 across the 3 of their papers we have counts for
3 papers
eess.AS2021★ 3 cited
Improved parallel WaveGAN vocoder with perceptually weighted spectrogram loss
Eunwoo Song, Ryuichi Yamamoto, Min-Jae Hwang +3
This paper proposes a spectral-domain perceptual weighting technique for Parallel WaveGAN-based text-to-speech (TTS) systems. The recently proposed Parallel WaveGAN vocoder success…
eess.AS2020
Neural text-to-speech with a modeling-by-generation excitation vocoder
Eunwoo Song, Min-Jae Hwang, Ryuichi Yamamoto +3
This paper proposes a modeling-by-generation (MbG) excitation vocoder for a neural text-to-speech (TTS) system. Recently proposed neural excitation vocoders can realize qualified w…
eess.AS2019★ 7 cited
Effective parameter estimation methods for an ExcitNet model in generative text-to-speech systems
Ohsung Kwon, Eunwoo Song, Jae-Min Kim +1
In this paper, we propose a high-quality generative text-to-speech (TTS) system using an effective spectrum and excitation estimation method. Our previous research verified the eff…