Showing cs.SDShow all
3 papers · 1 filter
cs.SD2024
How phonemes contribute to deep speaker models?
Pengqi Li, Tianhao Wang, Lantian Li +2
Which phonemes convey more speaker traits is a long-standing question, and various perception experiments were conducted with human subjects. For speaker recognition, studies were…
cs.SD2023
Visualizing data augmentation in deep speaker recognition
Pengqi Li, Lantian Li, Askar Hamdulla +1
Visualization is of great value in understanding the internal mechanisms of neural networks. Previous work found that LayerCAM is a reliable visualization tool for deep speaker mod…
cs.SD2022
Reliable Visualization for Deep Speaker Recognition
Pengqi Li, Lantian Li, Askar Hamdulla +1
In spite of the impressive success of convolutional neural networks (CNNs) in speaker recognition, our understanding to CNNs' internal functions is still limited. A major obstacle…