34 citations · 49 across the 8 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2022
Fusion of Self-supervised Learned Models for MOS Prediction
Zhengdong Yang, Wangjin Zhou, Chenhui Chu +4
We participated in the mean opinion score (MOS) prediction challenge, 2022. This challenge aims to predict MOS scores of synthetic speech on two tracks, the main track and a more c…
cs.SD2021★ 7 cited
MELONS: generating melody with long-term structure using transformers and structure graph
Yi Zou, Pei Zou, Yi Zhao +3
The creation of long melody sequences requires effective expression of coherent musical structure. However, there is no clear representation of musical structure. Recent works on m…
cs.SD2020★ 1 cited
Pretraining Strategies, Waveform Model Choice, and Acoustic Configurations for Multi-Speaker End-to-End Speech Synthesis
Erica Cooper, Xin Wang, Yi Zhao +2
We explore pretraining strategies including choice of base corpus with the aim of choosing the best strategy for zero-shot multi-speaker end-to-end synthesis. We also examine choic…