4 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.CL2023★ 4 cited
Unsupervised Data Selection for TTS: Using Arabic Broadcast News as a Case Study
Massa Baali, Tomoki Hayashi, Hamdy Mubarak +4
Several high-resource Text to Speech (TTS) systems currently produce natural, well-established human-like speech. In contrast, low-resource languages, including Arabic, have very l…
eess.AS2022
ESPnet-ONNX: Bridging a Gap Between Research and Production
Masao Someki, Yosuke Higuchi, Tomoki Hayashi +1
In the field of deep learning, researchers often focus on inventing novel neural network models and improving benchmarks. In contrast, application developers are interested in maki…
cs.SD2022
Muskits: an End-to-End Music Processing Toolkit for Singing Voice Synthesis
Jiatong Shi, Shuai Guo, Tao Qian +9
This paper introduces a new open-source platform named Muskits for end-to-end music processing, which mainly focuses on end-to-end singing voice synthesis (E2E-SVS). Muskits suppor…