5 citations · 5 across the 4 of their papers we have counts for
5 papers
Attention-based multi-task learning for speech-enhancement and speaker-identification in multi-speaker dialogue scenario
Chiang-Jen Peng, Yun-Ju Chan, Cheng Yu +3
Multi-task learning (MTL) and attention mechanism have been proven to effectively extract robust acoustic features for various speech-related tasks in noisy environments. In this s…
MoEVC: A Mixture-of-experts Voice Conversion System with Sparse Gating Mechanism for Accelerating Online Computation
Yu-Tao Chang, Yuan-Hong Yang, Yu-Huai Peng +4
With the recent advancements of deep learning technologies, the performance of voice conversion (VC) in terms of quality and similarity has been significantly improved. However, he…
Reinforcement Learning Based Speech Enhancement for Robust Speech Recognition
Yih-Liang Shen, Chao-Yuan Huang, Syu-Siang Wang +3
Conventional deep neural network (DNN)-based speech enhancement (SE) approaches aim to minimize the mean square error (MSE) between enhanced speech and clean reference. The MSE-opt…
Singing voice correction using canonical time warping
Yin-Jyun Luo, Ming-Tso Chen, Tai-Shih Chi +1
Expressive singing voice correction is an appealing but challenging problem. A robust time-warping algorithm which synchronizes two singing recordings can provide a promising solut…
Neural Network Based Next-Song Recommendation
Kai-Chun Hsu, Szu-Yu Chou, Yi-Hsuan Yang +1
Recently, the next-item/basket recommendation system, which considers the sequential relation between bought items, has drawn attention of researchers. The utilization of sequentia…