67 citations · 155 across the 17 of their papers we have counts for
6 papers · 1 filter
ForkNet: Simultaneous Time and Time-Frequency Domain Modeling for Speech Enhancement
Feng Dang, Qi Hu, Pengyuan Zhang +1
Previous research in speech enhancement has mostly focused on modeling time or time-frequency domain information alone, with little consideration given to the potential benefits of…
The HCCL Speaker Verification System for Far-Field Speaker Verification Challenge
Zhuo Li, Ce Fang, Runqiu Xiao +3
This paper describes the systems submitted by team HCCL to the Far-Field Speaker Verification Challenge. Our previous work in the AIshell Speaker Verification Challenge 2019 shows…
A Model Compression Method with Matrix Product Operators for Speech Enhancement
Xingwei Sun, Ze-Feng Gao, Zhong-Yi Lu +2
The deep neural network (DNN) based speech enhancement approaches have achieved promising performance. However, the number of parameters involved in these methods is usually enormo…
Stream Attention for far-field multi-microphone ASR
Xiaofei Wang, Yonghong Yan, Hynek Hermansky
A stream attention framework has been applied to the posterior probabilities of the deep neural network (DNN) to improve the far-field automatic speech recognition (ASR) performanc…
Relative Transfer Function Inverse Regression from Low Dimensional Manifold
Ziteng Wang, Emmanuel Vincent, Yonghong Yan
In room acoustic environments, the Relative Transfer Functions (RTFs) are controlled by few underlying modes of variability. Accordingly, they are confined to a low-dimensional man…
Rank-1 Constrained Multichannel Wiener Filter for Speech Recognition in Noisy Environments
Ziteng Wang, Emmanuel Vincent, Romain Serizel +1
Multichannel linear filters, such as the Multichannel Wiener Filter (MWF) and the Generalized Eigenvalue (GEV) beamformer are popular signal processing techniques which can improve…