13 citations · 24 across the 10 of their papers we have counts for
6 papers · 1 filter
DCCRN-KWS: an audio bias based model for noise robust small-footprint keyword spotting
Shubo Lv, Xiong Wang, Sining Sun +2
Real-world complex acoustic environments especially the ones with a low signal-to-noise ratio (SNR) will bring tremendous challenges to a keyword spotting (KWS) system. Inspired by…
A practical framework for multi-domain speech recognition and an instance sampling method to neural language modeling
Yike Zhang, Xiaobing Feng, Yi Liu +2
Automatic speech recognition (ASR) systems used on smart phones or vehicles are usually required to process speech queries from very different domains. In such situations, a vanill…
Improving Accent Identification and Accented Speech Recognition Under a Framework of Self-supervised Learning
Keqi Deng, Songjun Cao, Long Ma
Recently, self-supervised pre-training has gained success in automatic speech recognition (ASR). However, considering the difference between speech accents in real scenarios, how t…
Improving Streaming Transformer Based ASR Under a Framework of Self-supervised Learning
Songjun Cao, Yueteng Kang, Yanzhe Fu +4
Recently self-supervised learning has emerged as an effective approach to improve the performance of automatic speech recognition (ASR). Under such a framework, the neural network…
Improving Speech Recognition Accuracy of Local POI Using Geographical Models
Songjun Cao, Yike Zhang, Xiaobing Feng +1
Nowadays voice search for points of interest (POI) is becoming increasingly popular. However, speech recognition for local POI has remained to be a challenge due to multi-dialect a…
Tiny Transducer: A Highly-efficient Speech Recognition Model on Edge Devices
Yuekai Zhang, Sining Sun, Long Ma
This paper proposes an extremely lightweight phone-based transducer model with a tiny decoding graph on edge devices. First, a phone synchronous decoding (PSD) algorithm based on b…