activity
20182022
most citedGenerative adversarial network-based glottal waveform model for statistical parametric speech synthesis

50 citations · 58 across the 4 of their papers we have counts for

collaborators
Showing eess.ASShow all

7 papers · 1 filter

eess.AS20227 cited

Formant Tracking Using Quasi-Closed Phase Forward-Backward Linear Prediction Analysis and Deep Neural Networks

Dhananjaya Gowda, Bajibabu Bollepalli, Sudarsana Reddy Kadiri +1

Formant tracking is investigated in this study by using trackers based on dynamic programming (DP) and deep neural nets (DNNs). Using the DP approach, six formant estimation method…

eess.AS2019

ASVspoof 2019: A large-scale public database of synthesized, converted and replayed speech

Xin Wang, Junichi Yamagishi, Massimiliano Todisco +37

Automatic speaker verification (ASV) is one of the most natural and convenient means of biometric person recognition. Unfortunately, just like all other biometric systems, ASV is v…

eess.AS20191 cited

GELP: GAN-Excited Linear Prediction for Speech Synthesis from Mel-spectrogram

Lauri Juvela, Bajibabu Bollepalli, Junichi Yamagishi +1

Recent advances in neural network -based text-to-speech have reached human level naturalness in synthetic speech. The present sequence-to-sequence models can directly map text to m…

eess.AS201950 cited

Generative adversarial network-based glottal waveform model for statistical parametric speech synthesis

Bajibabu Bollepalli, Lauri Juvela, Paavo Alku

Recent studies have shown that text-to-speech synthesis quality can be improved by using glottal vocoding. This refers to vocoders that parameterize speech into two parts, the glot…

eess.AS2018

Waveform generation for text-to-speech synthesis using pitch-synchronous multi-scale generative adversarial networks

Lauri Juvela, Bajibabu Bollepalli, Junichi Yamagishi +1

The state-of-the-art in text-to-speech synthesis has recently improved considerably due to novel neural waveform generation methods, such as WaveNet. However, these methods suffer…

eess.AS2018

Speaker-independent raw waveform model for glottal excitation

Lauri Juvela, Vassilis Tsiaras, Bajibabu Bollepalli +3

Recent speech technology research has seen a growing interest in using WaveNets as statistical vocoders, i.e., generating speech waveforms from acoustic features. These models have…