24 citations · 90 across the 18 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2020★ 10 cited
Phonetic Posteriorgrams based Many-to-Many Singing Voice Conversion via Adversarial Training
Haohan Guo, Heng Lu, Na Hu +5
This paper describes an end-to-end adversarial singing voice conversion (EA-SVC) approach. It can directly generate arbitrary singing waveform by given phonetic posteriorgram (PPG)…
cs.SD2020
Improving RNN Transducer With Target Speaker Extraction and Neural Uncertainty Estimation
Jiatong Shi, Chunlei Zhang, Chao Weng +3
Target-speaker speech recognition aims to recognize target-speaker speech from noisy environments with background noise and interfering speakers. This work presents a joint framewo…