20 citations · 22 across the 2 of their papers we have counts for
2 papers
eess.AS2021★ 20 cited
An Encoder-Decoder Based Audio Captioning System With Transfer and Reinforcement Learning
Xinhao Mei, Qiushi Huang, Xubo Liu +10
Automated audio captioning aims to use natural language to describe the content of audio data. This paper presents an audio captioning system with an encoder-decoder architecture,…
cs.SD2020★ 2 cited
A Principle Solution for Enroll-Test Mismatch in Speaker Recognition
Lantian Li, Dong Wang, Jiawen Kang +4
Mismatch between enrollment and test conditions causes serious performance degradation on speaker recognition systems. This paper presents a statistics decomposition (SD) approach…