Showing eess.ASShow all
2 papers · 1 filter
eess.AS2024
SELM: Speech Enhancement Using Discrete Tokens and Language Models
Ziqian Wang, Xinfa Zhu, Zihan Zhang +4
Language models (LMs) have shown superior performances in various speech generation tasks recently, demonstrating their powerful ability for semantic context modeling. Given the in…
eess.AS2023
Boosting Multi-Speaker Expressive Speech Synthesis with Semi-supervised Contrastive Learning
Xinfa Zhu, Yuke Li, Yi Lei +3
This paper aims to build a multi-speaker expressive TTS system, synthesizing a target speaker's speech with multiple styles and emotions. To this end, we propose a novel contrastiv…