23 citations · 28 across the 5 of their papers we have counts for
5 papers
A Cyclical Approach to Synthetic and Natural Speech Mismatch Refinement of Neural Post-filter for Low-cost Text-to-speech System
Yi-Chiao Wu, Patrick Lumban Tobing, Kazuki Yasuhara +3
Neural-based text-to-speech (TTS) systems achieve very high-fidelity speech generation because of the rapid neural network developments. However, the huge labeled corpus and high c…
Unified Source-Filter GAN with Harmonic-plus-Noise Source Excitation Generation
Reo Yoneyama, Yi-Chiao Wu, Tomoki Toda
This paper introduces a unified source-filter network with a harmonic-plus-noise source excitation generation mechanism. In our previous work, we proposed unified Source-Filter GAN…
Non-parallel Voice Conversion System with WaveNet Vocoder and Collapsed Speech Suppression
Yi-Chiao Wu, Patrick Lumban Tobing, Kazuhiro Kobayashi +2
In this paper, we integrate a simple non-parallel voice conversion (VC) system with a WaveNet (WN) vocoder and a proposed collapsed speech suppression technique. The effectiveness…
Voice Conversion from Non-parallel Corpora Using Variational Auto-encoder
Chin-Cheng Hsu, Hsin-Te Hwang, Yi-Chiao Wu +2
We propose a flexible framework for spectral conversion (SC) that facilitates training with unaligned corpora. Many SC frameworks require parallel corpora, phonetic alignments, or…
Dictionary Update for NMF-based Voice Conversion Using an Encoder-Decoder Network
Chin-Cheng Hsu, Hsin-Te Hwang, Yi-Chiao Wu +2
In this paper, we propose a dictionary update method for Nonnegative Matrix Factorization (NMF) with high dimensional data in a spectral conversion (SC) task. Voice conversion has…