2 citations · 4 across the 4 of their papers we have counts for
4 papers
An empirical study of weakly supervised audio tagging embeddings for general audio representations
Heinrich Dinkel, Zhiyong Yan, Yongqing Wang +2
We study the usability of pre-trained weakly supervised audio tagging (AT) models as feature extractors for general audio representations. We mainly analyze the feasibility of tran…
UniKW-AT: Unified Keyword Spotting and Audio Tagging
Heinrich Dinkel, Yongqing Wang, Zhiyong Yan +2
Within the audio research community and the industry, keyword spotting (KWS) and audio tagging (AT) are seen as two distinct tasks and research fields. However, from a technical po…
Pseudo strong labels for large scale weakly supervised audio tagging
Heinrich Dinkel, Zhiyong Yan, Yongqing Wang +2
Large-scale audio tagging datasets inevitably contain imperfect labels, such as clip-wise annotated (temporally weak) tags with no exact on- and offsets, due to a high manual label…
speechocean762: An Open-Source Non-native English Speech Corpus For Pronunciation Assessment
Junbo Zhang, Zhiwen Zhang, Yongqing Wang +6
This paper introduces a new open-source speech corpus named "speechocean762" designed for pronunciation assessment use, consisting of 5000 English utterances from 250 non-native sp…