6 citations · 21 across the 24 of their papers we have counts for
4 papers · 1 filter
Extracting speaker and emotion information from self-supervised speech models via channel-wise correlations
Themos Stafylakis, Ladislav Mosner, Sofoklis Kakouros +3
Self-supervised learning of speech representations from large amounts of unlabeled data has enabled state-of-the-art results in several speech processing tasks. Aggregating these s…
An attention-based backend allowing efficient fine-tuning of transformer models for speaker verification
Junyi Peng, Oldrich Plchot, Themos Stafylakis +3
In recent years, self-supervised learning paradigm has received extensive attention due to its great success in various down-stream tasks. However, the fine-tuning strategies for a…
Speaker adaptation for Wav2vec2 based dysarthric ASR
Murali Karthick Baskar, Tim Herzig, Diana Nguyen +4
Dysarthric speech recognition has posed major challenges due to lack of training data and heavy mismatch in speaker characteristics. Recent ASR systems have benefited from readily…
DPCCN: Densely-Connected Pyramid Complex Convolutional Network for Robust Speech Separation And Extraction
Jiangyu Han, Yanhua Long, Lukas Burget +1
In recent years, a number of time-domain speech separation methods have been proposed. However, most of them are very sensitive to the environments and wide domain coverage tasks.…