2 papers
cs.SD2026
FAC-FACodec: Controllable Zero-Shot Foreign Accent Conversion with Factorized Speech Codec
Yurii Halychanskyi, Cameron Churchwell, Yutong Wen +1
Previous accent conversion (AC) methods, including foreign accent conversion (FAC), lack explicit control over the degree of modification. Because accent modification can alter the…
cs.SD2026
Audio Generation Through Score-Based Generative Modeling: Design Principles and Implementation
Ge Zhu, Yutong Wen, Zhiyao Duan
Diffusion models have emerged as powerful deep generative techniques, producing high-quality and diverse samples in applications in various domains including audio. While existing…