13 citations · 84 across the 55 of their papers we have counts for
6 papers · 2 filters
An Investigation of Distribution Alignment in Multi-Genre Speaker Recognition
Zhenyu Zhou, Junhui Chen, Namin Wang +2
Multi-genre speaker recognition is becoming increasingly popular due to its ability to better represent the complexities of real-world applications. However, a major challenge is t…
Multi-Domain Adaptation by Self-Supervised Learning for Speaker Verification
Wan Lin, Lantian Li, Dong Wang
In real-world applications, speaker recognition models often face various domain-mismatch challenges, leading to a significant drop in performance. Although numerous domain adaptat…
Spot keywords from very noisy and mixed speech
Ying Shi, Dong Wang, Lantian Li +2
Most existing keyword spotting research focuses on conditions with slight or moderate noise. In this paper, we try to tackle a more challenging task: detecting keywords buried unde…
A Multi-Scale Attentive Transformer for Multi-Instrument Symbolic Music Generation
Xipin Wei, Junhui Chen, Zirui Zheng +3
Recently, multi-instrument music generation has become a hot topic. Different from single-instrument generation, multi-instrument generation needs to consider inter-track harmony b…
Visualizing data augmentation in deep speaker recognition
Pengqi Li, Lantian Li, Askar Hamdulla +1
Visualization is of great value in understanding the internal mechanisms of neural networks. Previous work found that LayerCAM is a reliable visualization tool for deep speaker mod…
Ordered and Binary Speaker Embedding
Jiaying Wang, Xianglong Wang, Namin Wang +2
Modern speaker recognition systems represent utterances by embedding vectors. Conventional embedding vectors are dense and non-structural. In this paper, we propose an ordered bina…