7 citations · 7 across the 1 of their papers we have counts for
2 papers
eess.AS2020
Sequence-to-sequence Singing Voice Synthesis with Perceptual Entropy Loss
Jiatong Shi, Shuai Guo, Nan Huo +2
The neural network (NN) based singing voice synthesis (SVS) systems require sufficient data to train well and are prone to over-fitting due to data scarcity. However, we often enco…
eess.AS2020★ 7 cited
Context-aware Goodness of Pronunciation for Computer-Assisted Pronunciation Training
Jiatong Shi, Nan Huo, Qin Jin
Mispronunciation detection is an essential component of the Computer-Assisted Pronunciation Training (CAPT) systems. State-of-the-art mispronunciation detection models use Deep Neu…