papers

Publications (17)

eess.AS2023

DCCRN-KWS: an audio bias based model for noise robust small-footprint keyword spotting

Shubo Lv, Xiong Wang, Sining Sun +2

Real-world complex acoustic environments especially the ones with a low signal-to-noise ratio (SNR) will bring tremendous challenges to a keyword spotting (KWS) system. Inspired by…

cs.CL2024

Skipformer: A Skip-and-Recover Strategy for Efficient Speech Recognition

Wenjing Zhu, Sining Sun, Changhao Shan +2

Conformer-based attention models have become the de facto backbone model for Automatic Speech Recognition tasks. A blank symbol is usually introduced to align the input and output…

eess.AS2021

Improving Streaming Transformer Based ASR Under a Framework of Self-supervised Learning

Songjun Cao, Yueteng Kang, Yanzhe Fu +4

Recently self-supervised learning has emerged as an effective approach to improve the performance of automatic speech recognition (ASR). Under such a framework, the neural network…

eess.AS2022

Leveraging Acoustic Contextual Representation by Audio-textual Cross-modal Learning for Conversational ASR

Kun Wei, Yike Zhang, Sining Sun +2

Leveraging context information is an intuitive idea to improve performance on conversational automatic speech recognition(ASR). Previous works usually adopt recognized hypotheses o…

cs.CL2018

Training Augmentation with Adversarial Examples for Robust Speech Recognition

Sining Sun, Ching-Feng Yeh, Mari Ostendorf +2

This paper explores the use of adversarial examples in training speech recognition systems to increase robustness of deep neural network acoustic models. During training, the fast…

cs.SD2018

Investigating Generative Adversarial Networks based Speech Dereverberation for Robust Speech Recognition

Ke Wang, Junbo Zhang, Sining Sun +3

We investigate the use of generative adversarial networks (GANs) in speech dereverberation for robust speech recognition. GANs have been recently studied for speech enhancement to…