8 citations · 11 across the 6 of their papers we have counts for
5 papers
PadAug: Robust Speaker Verification with Simple Waveform-Level Silence Padding
Zijun Huang, Chengdong Liang, Jiadi Yao +1
The presence of non-speech segments in utterances often leads to the performance degradation of speaker verification. Existing systems usually use voice activation detection as a p…
Fast-U2++: Fast and Accurate End-to-End Speech Recognition in Joint CTC/Attention Frames
Chengdong Liang, Xiao-Lei Zhang, BinBin Zhang +5
Recently, the unified streaming and non-streaming two-pass (U2/U2++) end-to-end model for speech recognition has shown great performance in terms of streaming capability, accuracy…
WeKws: A production first small-footprint end-to-end Keyword Spotting Toolkit
Jie Wang, Menglong Xu, Jingyong Hou +4
Keyword spotting (KWS) enables speech-based user interaction and gradually becomes an indispensable component of smart devices. Recently, end-to-end (E2E) methods have become the m…
End-to-end Two-dimensional Sound Source Localization With Ad-hoc Microphone Arrays
Yijun Gong, Shupei Liu, Xiao-Lei Zhang
Conventional sound source localization methods are mostly based on a single microphone array that consists of multiple microphones. They are usually formulated as the estimation of…
Libri-adhoc40: A dataset collected from synchronized ad-hoc microphone arrays
Shanzheng Guan, Shupei Liu, Junqi Chen +8
Recently, there is a research trend on ad-hoc microphone arrays. However, most research was conducted on simulated data. Although some data sets were collected with a small number…