activity
20202026
most citedUnidirectional Memory-Self-Attention Transducer for Online Speech Recognition

2 citations · 5 across the 10 of their papers we have counts for

collaborators
Showing eess.ASShow all

6 papers · 1 filter

eess.AS2021

Loss Prediction: End-to-End Active Learning Approach For Speech Recognition

Jian Luo, Jianzong Wang, Ning Cheng +1

End-to-end speech recognition systems usually require huge amounts of labeling resource, while annotating the speech data is complicated and expensive. Active learning is the solut…

eess.AS2021

Dropout Regularization for Self-Supervised Learning of Transformer Encoder Speech Representation

Jian Luo, Jianzong Wang, Ning Cheng +1

Predicting the altered acoustic frames is an effective way of self-supervised learning for speech representation. However, it is challenging to prevent the pretrained model from ov…

eess.AS20212 cited

Unidirectional Memory-Self-Attention Transducer for Online Speech Recognition

Jian Luo, Jianzong Wang, Ning Cheng +1

Self-attention models have been successfully applied in end-to-end speech recognition systems, which greatly improve the performance of recognition accuracy. However, such attentio…

eess.AS20201 cited

Multi-QuartzNet: Multi-Resolution Convolution for Speech Recognition with Multi-Layer Feature Fusion

Jian Luo, Jianzong Wang, Ning Cheng +2

In this paper, we propose an end-to-end speech recognition network based on Nvidia's previous QuartzNet model. We try to promote the model performance, and design three components:…

eess.AS2020

End-to-end Silent Speech Recognition with Acoustic Sensing

Jian Luo, Jianzong Wang, Ning Cheng +2

Silent speech interfaces (SSI) has been an exciting area of recent interest. In this paper, we present a non-invasive silent speech interface that uses inaudible acoustic signals t…

eess.AS20202 cited

MLNET: An Adaptive Multiple Receptive-field Attention Neural Network for Voice Activity Detection

Zhenpeng Zheng, Jianzong Wang, Ning Cheng +2

Voice activity detection (VAD) makes a distinction between speech and non-speech and its performance is of crucial importance for speech based services. Recently, deep neural netwo…