activity
20182023
most citedSeamlessM4T: Massively Multilingual & Multimodal Machine Translation

14 citations · 23 across the 8 of their papers we have counts for

collaborators
Showing eess.ASShow all

8 papers · 1 filter

eess.AS2022

TTS-by-TTS 2: Data-selective augmentation for neural speech synthesis using ranking support vector machine with variational autoencoder

Eunwoo Song, Ryuichi Yamamoto, Ohsung Kwon +6

Recent advances in synthetic speech quality have enabled us to train text-to-speech (TTS) systems by using synthetic corpora. However, merely increasing the amount of synthetic dat…

eess.AS2021★ 3 cited

Improved parallel WaveGAN vocoder with perceptually weighted spectrogram loss

Eunwoo Song, Ryuichi Yamamoto, Min-Jae Hwang +3

This paper proposes a spectral-domain perceptual weighting technique for Parallel WaveGAN-based text-to-speech (TTS) systems. The recently proposed Parallel WaveGAN vocoder success…

eess.AS2020★ 5 cited

TTS-by-TTS: TTS-driven Data Augmentation for Fast and High-Quality Speech Synthesis

Min-Jae Hwang, Ryuichi Yamamoto, Eunwoo Song +1

In this paper, we propose a text-to-speech (TTS)-driven data augmentation method for improving the quality of a non-autoregressive (AR) TTS system. Recently proposed non-AR models,…

eess.AS2020

Parallel waveform synthesis based on generative adversarial networks with voicing-aware conditional discriminators

Ryuichi Yamamoto, Eunwoo Song, Min-Jae Hwang +1

This paper proposes voicing-aware conditional discriminators for Parallel WaveGAN-based waveform synthesis systems. In this framework, we adopt a projection-based conditioning meth…

eess.AS2020

Neural text-to-speech with a modeling-by-generation excitation vocoder

Eunwoo Song, Min-Jae Hwang, Ryuichi Yamamoto +3

This paper proposes a modeling-by-generation (MbG) excitation vocoder for a neural text-to-speech (TTS) system. Recently proposed neural excitation vocoders can realize qualified w…

eess.AS2020

Improving LPCNet-based Text-to-Speech with Linear Prediction-structured Mixture Density Network

Min-Jae Hwang, Eunwoo Song, Ryuichi Yamamoto +2

In this paper, we propose an improved LPCNet vocoder using a linear prediction (LP)-structured mixture density network (MDN). The recently proposed LPCNet vocoder has successfully…