activity
20152023
most citedWaveform Modeling and Generation Using Hierarchical Recurrent Neural Networks for Speech Bandwidth Extension

65 citations · 295 across the 29 of their papers we have counts for

collaborators
Showing cs.SDShow all

7 papers · 1 filter

cs.SD2022★ 2 cited

A Complementary Joint Training Approach Using Unpaired Speech and Text for Low-Resource Automatic Speech Recognition

Ye-Qian Du, Jie Zhang, Qiu-Shi Zhu +4

Unpaired data has shown to be beneficial for low-resource automatic speech recognition~(ASR), which can be involved in the design of hybrid models with multi-task training or langu…

cs.SD2020

Correlating Subword Articulation with Lip Shapes for Embedding Aware Audio-Visual Speech Enhancement

Hang Chen, Jun Du, Yu Hu +3

In this paper, we propose a visual embedding approach to improving embedding aware speech enhancement (EASE) by synchronizing visual lip frames at the phone and place of articulati…

cs.SD2019★ 8 cited

Singing Voice Synthesis Using Deep Autoregressive Neural Networks for Acoustic Modeling

Yuan-Hao Yi, Yang Ai, Zhen-Hua Ling +1

This paper presents a method of using autoregressive neural networks for the acoustic modeling of singing voice synthesis (SVS). Singing voice differs from speech and it contains m…

cs.SD2018

Improving Sequence-to-Sequence Acoustic Modeling by Adding Text-Supervision

Jing-Xuan Zhang, Zhen-Hua Ling, Yuan Jiang +3

This paper presents methods of making using of text supervision to improve the performance of sequence-to-sequence (seq2seq) voice conversion. Compared with conventional frame-to-f…

cs.SD2018

Sequence-to-Sequence Acoustic Modeling for Voice Conversion

Jing-Xuan Zhang, Zhen-Hua Ling, Li-Juan Liu +2

In this paper, a neural network named Sequence-to-sequence ConvErsion NeTwork (SCENT) is presented for acoustic modeling in voice conversion. At training stage, a SCENT model is es…

cs.SD2018★ 65 cited

Waveform Modeling and Generation Using Hierarchical Recurrent Neural Networks for Speech Bandwidth Extension

Zhen-Hua Ling, Yang Ai, Yu Gu +1

This paper presents a waveform modeling and generation method using hierarchical recurrent neural networks (HRNN) for speech bandwidth extension (BWE). Different from conventional…