1 paper
Seung-won Park, Doo-young Kim, Myun-chul Joe
We propose Cotatron, a transcription-guided speech encoder for speaker-independent linguistic representation. Cotatron is based on the multispeaker TTS architecture and can be trai…