3 citations · 4 across the 3 of their papers we have counts for
3 papers
cs.SD2024★ 1 cited
Whisper-PMFA: Partial Multi-Scale Feature Aggregation for Speaker Verification using Whisper Models
Yiyang Zhao, Shuai Wang, Guangzhi Sun +4
In this paper, Whisper, a large-scale pre-trained model for automatic speech recognition, is proposed to apply to speaker verification. A partial multi-scale feature aggregation (P…
cs.SD2023
Enhancing Quantised End-to-End ASR Models via Personalisation
Qiuming Zhao, Guangzhi Sun, Chao Zhang +2
Recent end-to-end automatic speech recognition (ASR) models have become increasingly larger, making them particularly challenging to be deployed on resource-constrained devices. Mo…
cs.LG2016★ 3 cited
Study on Feature Subspace of Archetypal Emotions for Speech Emotion Recognition
Xi Ma, Zhiyong Wu, Jia Jia +3
Feature subspace selection is an important part in speech emotion recognition. Most of the studies are devoted to finding a feature subspace for representing all emotions. However,…