6 citations · 8 across the 2 of their papers we have counts for
3 papers
cs.SD2020★ 6 cited
T-vectors: Weakly Supervised Speaker Identification Using Hierarchical Transformer Model
Yanpei Shi, Mingjie Chen, Qiang Huang +1
Identifying multiple speakers without knowing where a speaker's voice is in a recording is a challenging task. This paper proposes a hierarchical network with transformer encoders…
cs.SD2020
Towards Low-Resource StarGAN Voice Conversion using Weight Adaptive Instance Normalization
Mingjie Chen, Yanpei Shi, Thomas Hain
Many-to-many voice conversion with non-parallel training data has seen significant progress in recent years. StarGAN-based models have been interests of voice conversion. However,…
eess.AS2020★ 2 cited
Unsupervised Acoustic Unit Representation Learning for Voice Conversion using WaveNet Auto-encoders
Mingjie Chen, Thomas Hain
Unsupervised representation learning of speech has been of keen interest in recent years, which is for example evident in the wide interest of the ZeroSpeech challenges. This work…