1 citations · 2 across the 4 of their papers we have counts for
4 papers
A Multi-Resolution Front-End for End-to-End Speech Anti-Spoofing
Wei Liu, Meng Sun, Xiongwei Zhang +2
The choice of an optimal time-frequency resolution is usually a difficult but important step in tasks involving speech signal classification, e.g., speech anti-spoofing. The variat…
Low Bit-Rate Wideband Speech Coding: A Deep Generative Model based Approach
Gang Min, Xiongwei Zhang, Xia Zou +1
Traditional low bit-rate speech coding approach only handles narrowband speech at 8kHz, which limits further improvements in speech quality. Motivated by recent successful explorat…
When Automatic Voice Disguise Meets Automatic Speaker Verification
Linlin Zheng, Jiakang Li, Meng Sun +2
The technique of transforming voices in order to hide the real identity of a speaker is called voice disguise, among which automatic voice disguise (AVD) by modifying the spectral…
Deep Vocoder: Low Bit Rate Compression of Speech with Deep Autoencoder
Gang Min, Changqing Zhang, Xiongwei Zhang +1
Inspired by the success of deep neural networks (DNNs) in speech processing, this paper presents Deep Vocoder, a direct end-to-end low bit rate speech compression method with deep…