1 citations · 1 across the 6 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2024
The X-LANCE Technical Report for Interspeech 2024 Speech Processing Using Discrete Speech Unit Challenge
Yiwei Guo, Chenrun Wang, Yifan Yang +9
Discrete speech tokens have been more and more popular in multiple speech processing fields, including automatic speech recognition (ASR), text-to-speech (TTS) and singing voice sy…
eess.AS2023★ 1 cited
Expressive TTS Driven by Natural Language Prompts Using Few Human Annotations
Hanglei Zhang, Yiwei Guo, Sen Liu +2
Expressive text-to-speech (TTS) aims to synthesize speeches with human-like tones, moods, or even artistic attributes. Recent advancements in expressive TTS empower users with the…
eess.AS2023
DiffVoice: Text-to-Speech with Latent Diffusion
Zhijun Liu, Yiwei Guo, Kai Yu
In this work, we present DiffVoice, a novel text-to-speech model based on latent diffusion. We propose to first encode speech signals into a phoneme-rate latent representation with…