3 papers
cs.LG2025
Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis
Weiwei Lin, Chenghan He
We propose a novel autoregressive modeling approach for speech synthesis, combining a variational autoencoder (VAE) with a multi-modal latent space and an autoregressive model that…
cs.SD2024
VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis
Weiwei Lin, Chenhang He, Man-Wai Mak +2
Achieving nuanced and accurate emulation of human voice has been a longstanding goal in artificial intelligence. Although significant progress has been made in recent years, the ma…
eess.AS2023
Asymmetric Clean Segments-Guided Self-Supervised Learning for Robust Speaker Verification
Chong-Xin Gan, Man-Wai Mak, Weiwei Lin +1
Contrastive self-supervised learning (CSL) for speaker verification (SV) has drawn increasing interest recently due to its ability to exploit unlabeled data. Performing data augmen…