7 citations · 21 across the 6 of their papers we have counts for
6 papers
A Real-time Robot-based Auxiliary System for Risk Evaluation of COVID-19 Infection
Wenqi Wei, Jianzong Wang, Jiteng Ma +2
In this paper, we propose a real-time robot-based auxiliary system for risk evaluation of COVID-19 infection. It combines real-time speech recognition, temperature measurement, key…
Large-scale Transfer Learning for Low-resource Spoken Language Understanding
Xueli Jia, Jianzong Wang, Zhiyong Zhang +2
End-to-end Spoken Language Understanding (SLU) models are made increasingly large and complex to achieve the state-ofthe-art accuracy. However, the increased complexity of a model…
Prosody Learning Mechanism for Speech Synthesis System Without Text Length Limit
Zhen Zeng, Jianzong Wang, Ning Cheng +1
Recent neural speech synthesis systems have gradually focused on the control of prosody to improve the quality of synthesized speech, but they rarely consider the variability of pr…
MLNET: An Adaptive Multiple Receptive-field Attention Neural Network for Voice Activity Detection
Zhenpeng Zheng, Jianzong Wang, Ning Cheng +2
Voice activity detection (VAD) makes a distinction between speech and non-speech and its performance is of crucial importance for speech based services. Recently, deep neural netwo…
AlignTTS: Efficient Feed-Forward Text-to-Speech System without Explicit Alignment
Zhen Zeng, Jianzong Wang, Ning Cheng +2
Targeting at both high efficiency and performance, we propose AlignTTS to predict the mel-spectrum in parallel. AlignTTS is based on a Feed-Forward Transformer which generates mel-…
GraphTTS: graph-to-sequence modelling in neural text-to-speech
Aolan Sun, Jianzong Wang, Ning Cheng +3
This paper leverages the graph-to-sequence method in neural text-to-speech (GraphTTS), which maps the graph embedding of the input sequence to spectrograms. The graphical inputs co…