activity
20192021
most citedSpatial and spectral deep attention fusion for multi-channel speech separation using deep embedding features

8 citations · 21 across the 7 of their papers we have counts for

collaborators

10 papers

cs.LG20213 cited

MS-MDA: Multisource Marginal Distribution Adaptation for Cross-subject and Cross-session EEG Emotion Recognition

Hao Chen, Ming Jin, Zhunan Li +3

As an essential element for the diagnosis and rehabilitation of psychiatric disorders, the electroencephalogram (EEG) based emotion recognition has achieved significant progress du…

cs.SD2020

Deep Time Delay Neural Network for Speech Enhancement with Full Data Learning

Cunhang Fan, Bin Liu, Jianhua Tao +3

Recurrent neural networks (RNNs) have shown significant improvements in recent years for speech enhancement. However, the model complexity and inference time cost of RNNs are much…

cs.SD20203 cited

Gated Recurrent Fusion with Joint Training Framework for Robust End-to-End Speech Recognition

Cunhang Fan, Jiangyan Yi, Jianhua Tao +3

The joint training framework for speech enhancement and recognition methods have obtained quite good performances for robust end-to-end automatic speech recognition (ASR). However,…

cs.SD20204 cited

Dynamic Attention Based Generative Adversarial Network with Phase Post-Processing for Speech Enhancement

Andong Li, Chengshi Zheng, Renhua Peng +2

The generative adversarial networks (GANs) have facilitated the development of speech enhancement recently. Nevertheless, the performance advantage is still limited when compared w…

eess.AS20201 cited

Simultaneous Denoising and Dereverberation Using Deep Embedding Features

Cunhang Fan, Jianhua Tao, Bin Liu +2

Monaural speech dereverberation is a very challenging task because no spatial cues can be used. When the additive noises exist, this task becomes more challenging. In this paper, w…

cs.CL2020

Adversarial Transfer Learning for Punctuation Restoration

Jiangyan Yi, Jianhua Tao, Ye Bai +2

Previous studies demonstrate that word embeddings and part-of-speech (POS) tags are helpful for punctuation restoration tasks. However, two drawbacks still exist. One is that word…