9 citations · 12 across the 4 of their papers we have counts for
5 papers
Summary on the ISCSLP 2022 Chinese-English Code-Switching ASR Challenge
Shuhao Deng, Chengfei Li, Jinfeng Bai +6
Code-switching automatic speech recognition becomes one of the most challenging and the most valuable scenarios of automatic speech recognition, due to the code-switching phenomeno…
Full Attention Bidirectional Deep Learning Structure for Single Channel Speech Enhancement
Yuzi Yan, Wei-Qiang Zhang, Michael T. Johnson
As the cornerstone of other important technologies, such as speech recognition and speech synthesis, speech enhancement is a critical area in audio signal processing. In this paper…
AdaSpeech 3: Adaptive Text to Speech for Spontaneous Style
Yuzi Yan, Xu Tan, Bohan Li +6
While recent text to speech (TTS) models perform very well in synthesizing reading-style (e.g., audiobook) speech, it is still challenging to synthesize spontaneous-style speech (e…
DeepRapper: Neural Rap Generation with Rhyme and Rhythm Modeling
Lanqing Xue, Kaitao Song, Duocai Wu +5
Rap generation, which aims to produce lyrics and corresponding singing beats, needs to model both rhymes and rhythms. Previous works for rap generation focused on rhyming lyrics bu…
SAM-GCNN: A Gated Convolutional Neural Network with Segment-Level Attention Mechanism for Home Activity Monitoring
Yu-Han Shen, Ke-Xin He, Wei-Qiang Zhang
In this paper, we propose a method for home activity monitoring. We demonstrate our model on dataset of Detection and Classification of Acoustic Scenes and Events (DCASE) 2018 Chal…