Showing cs.SDShow all
3 papers · 1 filter
cs.SD2024
Few-Shot Keyword Spotting from Mixed Speech
Junming Yuan, Ying Shi, LanTian Li +2
Few-shot keyword spotting (KWS) aims to detect unknown keywords with limited training samples. A commonly used approach is the pre-training and fine-tuning framework. While effecti…
cs.SD2024
How phonemes contribute to deep speaker models?
Pengqi Li, Tianhao Wang, Lantian Li +2
Which phonemes convey more speaker traits is a long-standing question, and various perception experiments were conducted with human subjects. For speaker recognition, studies were…
cs.SD2023
Visualizing data augmentation in deep speaker recognition
Pengqi Li, Lantian Li, Askar Hamdulla +1
Visualization is of great value in understanding the internal mechanisms of neural networks. Previous work found that LayerCAM is a reliable visualization tool for deep speaker mod…