3 citations · 3 across the 3 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2024
The Effect of Batch Size on Contrastive Self-Supervised Speech Representation Learning
Nik Vaessen, David A. van Leeuwen
Foundation models in speech are often trained using many GPUs, which implicitly leads to large effective batch sizes. In this paper we study the effect of batch size on pre-trainin…
cs.SD2023
Towards multi-task learning of speech and speaker recognition
Nik Vaessen, David A. van Leeuwen
We study multi-task learning for two orthogonal speech technology tasks: speech and speaker recognition. We use wav2vec2 as a base architecture with two task-specific output heads.…