18 citations · 18 across the 3 of their papers we have counts for
5 papers
Caption Feature Space Regularization for Audio Captioning
Yiming Zhang, Hong Yu, Ruoyi Du +2
Audio captioning aims at describing the content of audio clips with human language. Due to the ambiguity of audio, different people may perceive the same audio differently, resulti…
Language Identification with Deep Bottleneck Features
Zhanyu Ma, Hong Yu
In this paper we proposed an end-to-end short utterances speech language identification(SLD) approach based on a Long Short Term Memory (LSTM) neural network which is special suita…
Histogram Transform-based Speaker Identification
Zhanyu Ma, Hong Yu
A novel text-independent speaker identification (SI) method is proposed. This method uses the Mel-frequency Cepstral coefficients (MFCCs) and the dynamic information among adjacent…
Adversarial Network Bottleneck Features for Noise Robust Speaker Verification
Hong Yu, Zheng-Hua Tan, Zhanyu Ma +1
In this paper, we propose a noise robust bottleneck feature representation which is generated by an adversarial network (AN). The AN includes two cascade connected networks, an enc…
DNN Filter Bank Cepstral Coefficients for Spoofing Detection
Hong Yu, Zheng-Hua Tan, Zhanyu Ma +1
With the development of speech synthesis techniques, automatic speaker verification systems face the serious challenge of spoofing attack. In order to improve the reliability of sp…