10 citations · 10 across the 5 of their papers we have counts for
5 papers · 1 filter
Sparsely Shared LoRA on Whisper for Child Speech Recognition
Wei Liu, Ying Qin, Zhiyuan Peng +1
Whisper is a powerful automatic speech recognition (ASR) model. Nevertheless, its zero-shot performance on low-resource speech requires further improvement. Child speech, as a repr…
A study on the efficacy of model pre-training in developing neural text-to-speech system
Guangyan Zhang, Yichong Leng, Daxin Tan +5
In the development of neural text-to-speech systems, model pre-training with a large amount of non-target speakers' data is a common approach. However, in terms of ultimately achie…
Exploiting Pre-Trained ASR Models for Alzheimer's Disease Recognition Through Spontaneous Speech
Ying Qin, Wei Liu, Zhiyuan Peng +4
Alzheimer's disease (AD) is a progressive neurodegenerative disease and recently attracts extensive attention worldwide. Speech technology is considered a promising solution for th…
Applying the Information Bottleneck Principle to Prosodic Representation Learning
Guangyan Zhang, Ying Qin, Daxin Tan +1
This paper describes a novel design of a neural network-based speech generation model for learning prosodic representation.The problem of representation learning is formulated acco…
An End-to-End Approach to Automatic Speech Assessment for Cantonese-speaking People with Aphasia
Ying Qin, Yuzhong Wu, Tan Lee +1
Conventional automatic assessment of pathological speech usually follows two main steps: (1) extraction of pathology-specific features; (2) classification or regression on extracte…