2 papers
eess.AS2020
Exploiting Deep Sentential Context for Expressive End-to-End Speech Synthesis
Fengyu Yang, Shan Yang, Qinghua Wu +2
Attention-based seq2seq text-to-speech systems, especially those use self-attention networks (SAN), have achieved state-of-art performance. But an expressive corpus with rich proso…
cs.SD2018
End-to-end Models with auditory attention in Multi-channel Keyword Spotting
Haitong Zhang, Junbo Zhang, Yujun Wang
In this paper, we propose an attention-based end-to-end model for multi-channel keyword spotting (KWS), which is trained to optimize the KWS result directly. As a result, our model…