3 citations · 3 across the 5 of their papers we have counts for
Showing eess.ASShow all
2 papers · 1 filter
eess.AS2024
Improving Robustness of Diffusion-Based Zero-Shot Speech Synthesis via Stable Formant Generation
Changjin Han, Seokgi Lee, Gyuhyeon Nam +1
Diffusion models have achieved remarkable success in text-to-speech (TTS), even in zero-shot scenarios. Recent efforts aim to address the trade-off between inference speed and soun…
eess.AS2021
Axial Residual Networks for CycleGAN-based Voice Conversion
Jaeseong You, Gyuhyeon Nam, Dalhyun Kim +1
We propose a novel architecture and improved training objectives for non-parallel voice conversion. Our proposed CycleGAN-based model performs a shape-preserving transformation dir…