67 citations · 216 across the 25 of their papers we have counts for
5 papers · 1 filter
End-to-End Speaker-Dependent Voice Activity Detection
Yefei Chen, Shuai Wang, Yanmin Qian +1
Voice activity detection (VAD) is an essential pre-processing step for tasks such as automatic speech recognition (ASR) and speaker recognition. A basic goal is to remove silent se…
Future Vector Enhanced LSTM Language Model for LVCSR
Qi Liu, Yanmin Qian, Kai Yu
Language models (LM) play an important role in large vocabulary continuous speech recognition (LVCSR). However, traditional language models only predict next single word with given…
Modular End-to-end Automatic Speech Recognition Framework for Acoustic-to-word Model
Qi Liu, Zhehuai Chen, Hao Li +3
End-to-end (E2E) systems have played a more and more important role in automatic speech recognition (ASR) and achieved great performance. However, E2E systems recognize output word…
End-to-end spoofing detection with raw waveform CLDNNs
Heinrich Dinkel, Nanxin Chen, Yanmin Qian +1
Albeit recent progress in speaker verification generates powerful models, malicious attacks in the form of spoofed speech, are generally not coped with. Recent results in ASVSpoof2…
Margin Matters: Towards More Discriminative Deep Neural Network Embeddings for Speaker Recognition
Xu Xiang, Shuai Wang, Houjun Huang +2
Recently, speaker embeddings extracted from a speaker discriminative deep neural network (DNN) yield better performance than the conventional methods such as i-vector. In most case…