3 papers
cs.SD2024
How phonemes contribute to deep speaker models?
Pengqi Li, Tianhao Wang, Lantian Li +2
Which phonemes convey more speaker traits is a long-standing question, and various perception experiments were conducted with human subjects. For speaker recognition, studies were…
cs.SD2023
Visualizing data augmentation in deep speaker recognition
Pengqi Li, Lantian Li, Askar Hamdulla +1
Visualization is of great value in understanding the internal mechanisms of neural networks. Previous work found that LayerCAM is a reliable visualization tool for deep speaker mod…
cs.CV2023
DASTSiam: Spatio-Temporal Fusion and Discriminative Augmentation for Improved Siamese Tracking
Yucheng Huang, Eksan Firkat, Ziwang Xiao +2
Tracking tasks based on deep neural networks have greatly improved with the emergence of Siamese trackers. However, the appearance of targets often changes during tracking, which c…