65 citations · 264 across the 28 of their papers we have counts for
44 papers
Decoupled Pronunciation and Prosody Modeling in Meta-Learning-Based Multilingual Speech Synthesis
Yukun Peng, Zhenhua Ling
This paper presents a method of decoupled pronunciation and prosody modeling to improve the performance of meta-learning-based multilingual speech synthesis. The baseline meta-lear…
Neural Grapheme-to-Phoneme Conversion with Pre-trained Grapheme Models
Lu Dong, Zhi-Qiang Guo, Chao-Hong Tan +3
Neural network models have achieved state-of-the-art performance on grapheme-to-phoneme (G2P) conversion. However, their performance relies on large-scale pronunciation dictionarie…
MPC-BERT: A Pre-Trained Language Model for Multi-Party Conversation Understanding
Jia-Chen Gu, Chongyang Tao, Zhen-Hua Ling +3
Recently, various neural models for multi-party conversation (MPC) have achieved impressive improvements on a variety of tasks such as addressee recognition, speaker identification…
Partner Matters! An Empirical Study on Fusing Personas for Personalized Response Selection in Retrieval-Based Chatbots
Jia-Chen Gu, Hui Liu, Zhen-Hua Ling +3
Persona can function as the prior knowledge for maintaining the consistency of dialogue systems. Most of previous studies adopted the self persona in dialogue whose response was ab…
Emotion-Regularized Conditional Variational Autoencoder for Emotional Response Generation
Yu-Ping Ruan, Zhen-Hua Ling
This paper presents an emotion-regularized conditional variational autoencoder (Emo-CVAE) model for generating emotional conversation responses. In conventional CVAE-based emotiona…
Learning to Retrieve Entity-Aware Knowledge and Generate Responses with Copy Mechanism for Task-Oriented Dialogue Systems
Chao-Hong Tan, Xiaoyu Yang, Zi'ou Zheng +7
Task-oriented conversational modeling with unstructured knowledge access, as track 1 of the 9th Dialogue System Technology Challenges (DSTC 9), requests to build a system to genera…