6 citations · 9 across the 6 of their papers we have counts for
Showing cs.SDShow all
3 papers · 1 filter
cs.SD2020★ 6 cited
T-vectors: Weakly Supervised Speaker Identification Using Hierarchical Transformer Model
Yanpei Shi, Mingjie Chen, Qiang Huang +1
Identifying multiple speakers without knowing where a speaker's voice is in a recording is a challenging task. This paper proposes a hierarchical network with transformer encoders…
cs.SD2020
Towards Low-Resource StarGAN Voice Conversion using Weight Adaptive Instance Normalization
Mingjie Chen, Yanpei Shi, Thomas Hain
Many-to-many voice conversion with non-parallel training data has seen significant progress in recent years. StarGAN-based models have been interests of voice conversion. However,…
cs.SD2020
Supervised Speaker Embedding De-Mixing in Two-Speaker Environment
Yanpei Shi, Thomas Hain
Separating different speaker properties from a multi-speaker environment is challenging. Instead of separating a two-speaker signal in signal space like speech source separation, a…