most citedEditSpeech: A Text Based Speech Editing System Using Partial Inference and Bidirectional Fusion

1 citations · 1 across the 3 of their papers we have counts for

collaborators

5 papers

eess.AS2021

A study on the efficacy of model pre-training in developing neural text-to-speech system

Guangyan Zhang, Yichong Leng, Daxin Tan +5

In the development of neural text-to-speech systems, model pre-training with a large amount of non-target speakers' data is a common approach. However, in terms of ultimately achie…

eess.AS2021

Applying the Information Bottleneck Principle to Prosodic Representation Learning

Guangyan Zhang, Ying Qin, Daxin Tan +1

This paper describes a novel design of a neural network-based speech generation model for learning prosodic representation.The problem of representation learning is formulated acco…

eess.AS20211 cited

EditSpeech: A Text Based Speech Editing System Using Partial Inference and Bidirectional Fusion

Daxin Tan, Liqun Deng, Yu Ting Yeung +3

This paper presents the design, implementation and evaluation of a speech editing system, named EditSpeech, which allows a user to perform deletion, insertion and replacement of wo…

eess.AS2021

CUHK-EE Voice Cloning System for ICASSP 2021 M2VoC Challenge

Daxin Tan, Hingpang Huang, Guangyan Zhang +1

This paper presents the CUHK-EE voice cloning system for ICASSP 2021 M2VoC challenge. The challenge provides two Mandarin speech corpora: the AIShell-3 corpus of 218 speakers with…

eess.AS2020

Fine-grained Style Modeling, Transfer and Prediction in Text-to-Speech Synthesis via Phone-Level Content-Style Disentanglement

Daxin Tan, Tan Lee

This paper presents a novel design of neural network system for fine-grained style modeling, transfer and prediction in expressive text-to-speech (TTS) synthesis. Fine-grained mode…