17 citations · 46 across the 28 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2025
Analyzing and Improving Speaker Similarity Assessment for Speech Synthesis
Marc-André Carbonneau, Benjamin van Niekerk, Hugo Seuté +3
Modeling voice identity is challenging due to its multifaceted nature. In generative speech systems, identity is often assessed using automatic speaker verification (ASV) embedding…
cs.SD2022
GAN You Hear Me? Reclaiming Unconditional Speech Synthesis from Diffusion Models
Matthew Baas, Herman Kamper
We propose AudioStyleGAN (ASGAN), a new generative adversarial network (GAN) for unconditional speech synthesis. As in the StyleGAN family of image synthesis models, ASGAN maps sam…