most citedSelf-Attention Transducers for End-to-End Speech Recognition

85 citations · 123 across the 9 of their papers we have counts for

collaborators

11 papers

eess.AS202010 cited

Spike-Triggered Non-Autoregressive Transformer for End-to-End Speech Recognition

Zhengkun Tian, Jiangyan Yi, Jianhua Tao +3

Non-autoregressive transformer models have achieved extremely fast inference speed and comparable performance with autoregressive sequence-to-sequence models in neural machine tran…

eess.AS20205 cited

Listen Attentively, and Spell Once: Whole Sentence Generation via a Non-Autoregressive Architecture for Low-Latency Speech Recognition

Ye Bai, Jiangyan Yi, Jianhua Tao +3

Although attention based end-to-end models have achieved promising performance in speech recognition, the multi-pass forward computation in beam-search increases inference time cos…

eess.AS20201 cited

Simultaneous Denoising and Dereverberation Using Deep Embedding Features

Cunhang Fan, Jianhua Tao, Bin Liu +2

Monaural speech dereverberation is a very challenging task because no spatial cues can be used. When the additive noises exist, this task becomes more challenging. In this paper, w…

cs.CL2020

Adversarial Transfer Learning for Punctuation Restoration

Jiangyan Yi, Jianhua Tao, Ye Bai +2

Previous studies demonstrate that word embeddings and part-of-speech (POS) tags are helpful for punctuation restoration tasks. However, two drawbacks still exist. One is that word…

eess.AS2020

Deep Attention Fusion Feature for Speech Separation with End-to-End Post-filter Method

Cunhang Fan, Jianhua Tao, Bin Liu +3

In this paper, we propose an end-to-end post-filter method with deep attention fusion features for monaural speaker-independent speech separation. At first, a time-frequency domain…

cs.CL20203 cited

Rnn-transducer with language bias for end-to-end Mandarin-English code-switching speech recognition

Shuai Zhang, Jiangyan Yi, Zhengkun Tian +2

Recently, language identity information has been utilized to improve the performance of end-to-end code-switching (CS) speech recognition. However, previous works use an additional…