13 citations · 17 across the 5 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2023★ 1 cited
DiCLET-TTS: Diffusion Model based Cross-lingual Emotion Transfer for Text-to-Speech -- A Study between English and Mandarin
Tao Li, Chenxu Hu, Jian Cong +5
While the performance of cross-lingual TTS based on monolingual corpora has been significantly improved recently, generating cross-lingual speech still suffers from the foreign acc…
cs.SD2023
Diff-Foley: Synchronized Video-to-Audio Synthesis with Latent Diffusion Models
Simian Luo, Chuanhao Yan, Chenxu Hu +1
The Video-to-Audio (V2A) model has recently gained attention for its practical application in generating audio directly from silent videos, particularly in video/film production. H…
cs.SD2020★ 3 cited
CVC: Contrastive Learning for Non-parallel Voice Conversion
Tingle Li, Yichen Liu, Chenxu Hu +1
Cycle consistent generative adversarial network (CycleGAN) and variational autoencoder (VAE) based models have gained popularity in non-parallel voice conversion recently. However,…