2 citations · 5 across the 10 of their papers we have counts for
6 papers · 1 filter
Loss Prediction: End-to-End Active Learning Approach For Speech Recognition
Jian Luo, Jianzong Wang, Ning Cheng +1
End-to-end speech recognition systems usually require huge amounts of labeling resource, while annotating the speech data is complicated and expensive. Active learning is the solut…
Dropout Regularization for Self-Supervised Learning of Transformer Encoder Speech Representation
Jian Luo, Jianzong Wang, Ning Cheng +1
Predicting the altered acoustic frames is an effective way of self-supervised learning for speech representation. However, it is challenging to prevent the pretrained model from ov…
Unidirectional Memory-Self-Attention Transducer for Online Speech Recognition
Jian Luo, Jianzong Wang, Ning Cheng +1
Self-attention models have been successfully applied in end-to-end speech recognition systems, which greatly improve the performance of recognition accuracy. However, such attentio…
Multi-QuartzNet: Multi-Resolution Convolution for Speech Recognition with Multi-Layer Feature Fusion
Jian Luo, Jianzong Wang, Ning Cheng +2
In this paper, we propose an end-to-end speech recognition network based on Nvidia's previous QuartzNet model. We try to promote the model performance, and design three components:…
End-to-end Silent Speech Recognition with Acoustic Sensing
Jian Luo, Jianzong Wang, Ning Cheng +2
Silent speech interfaces (SSI) has been an exciting area of recent interest. In this paper, we present a non-invasive silent speech interface that uses inaudible acoustic signals t…
MLNET: An Adaptive Multiple Receptive-field Attention Neural Network for Voice Activity Detection
Zhenpeng Zheng, Jianzong Wang, Ning Cheng +2
Voice activity detection (VAD) makes a distinction between speech and non-speech and its performance is of crucial importance for speech based services. Recently, deep neural netwo…