15 citations · 61 across the 12 of their papers we have counts for
12 papers · 1 filter
Spot keywords from very noisy and mixed speech
Ying Shi, Dong Wang, Lantian Li +2
Most existing keyword spotting research focuses on conditions with slight or moderate noise. In this paper, we try to tackle a more challenging task: detecting keywords buried unde…
Contrastive Regularization for Multimodal Emotion Recognition Using Audio and Text
Fan Qian, Jiqing Han
Speech emotion recognition is a challenge and an important step towards more natural human-computer interaction (HCI). The popular approach is multimodal emotion recognition based…
Can We Trust Deep Speech Prior?
Ying Shi, Haolin Chen, Zhiyuan Tang +3
Recently, speech enhancement (SE) based on deep speech prior has attracted much attention, such as the variational auto-encoder with non-negative matrix factorization (VAE-NMF) arc…
LaFurca: Iterative Refined Speech Separation Based on Context-Aware Dual-Path Parallel Bi-LSTM
Ziqiang Shi, Rujie Liu, Jiqing Han
Deep neural network with dual-path bi-directional long short-term memory (BiLSTM) block has been proved to be very effective in sequence modeling, especially in speech separation,…
Acoustic Scene Classification by Implicitly Identifying Distinct Sound Events
Hongwei Song, Jiqing Han, Shiwen Deng +1
In this paper, we propose a new strategy for acoustic scene classification (ASC) , namely recognizing acoustic scenes through identifying distinct sound events. This differs from e…
A Multi-Task Learning Framework for Overcoming the Catastrophic Forgetting in Automatic Speech Recognition
Jiabin Xue, Jiqing Han, Tieran Zheng +2
Recently, data-driven based Automatic Speech Recognition (ASR) systems have achieved state-of-the-art results. And transfer learning is often used when those existing systems are a…