7 citations · 13 across the 2 of their papers we have counts for
3 papers
cs.CV2021★ 6 cited
MEmoBERT: Pre-training Model with Prompt-based Learning for Multimodal Emotion Recognition
Jinming Zhao, Ruichen Li, Qin Jin +2
Multimodal emotion recognition study is hindered by the lack of labelled corpora in terms of scale and diversity, due to the high annotation cost and label ambiguity. In this paper…
eess.AS2020
Sequence-to-sequence Singing Voice Synthesis with Perceptual Entropy Loss
Jiatong Shi, Shuai Guo, Nan Huo +2
The neural network (NN) based singing voice synthesis (SVS) systems require sufficient data to train well and are prone to over-fitting due to data scarcity. However, we often enco…
eess.AS2020★ 7 cited
Context-aware Goodness of Pronunciation for Computer-Assisted Pronunciation Training
Jiatong Shi, Nan Huo, Qin Jin
Mispronunciation detection is an essential component of the Computer-Assisted Pronunciation Training (CAPT) systems. State-of-the-art mispronunciation detection models use Deep Neu…