217 citations · 399 across the 26 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2022
Caption Feature Space Regularization for Audio Captioning
Yiming Zhang, Hong Yu, Ruoyi Du +2
Audio captioning aims at describing the content of audio clips with human language. Due to the ambiguity of audio, different people may perceive the same audio differently, resulti…
cs.SD2017★ 18 cited
Adversarial Network Bottleneck Features for Noise Robust Speaker Verification
Hong Yu, Zheng-Hua Tan, Zhanyu Ma +1
In this paper, we propose a noise robust bottleneck feature representation which is generated by an adversarial network (AN). The AN includes two cascade connected networks, an enc…
cs.SD2017
DNN Filter Bank Cepstral Coefficients for Spoofing Detection
Hong Yu, Zheng-Hua Tan, Zhanyu Ma +1
With the development of speech synthesis techniques, automatic speaker verification systems face the serious challenge of spoofing attack. In order to improve the reliability of sp…