2 papers
eess.AS2025
RapFlow-TTS: Rapid and High-Fidelity Text-to-Speech with Improved Consistency Flow Matching
Hyun Joon Park, Jeongmin Liu, Jin Sob Kim +3
We introduce RapFlow-TTS, a rapid and high-fidelity TTS acoustic model that leverages velocity consistency constraints in flow matching (FM) training. Although ordinary differentia…
cs.SD2024
Training Universal Vocoders with Feature Smoothing-Based Augmentation Methods for High-Quality TTS Systems
Jeongmin Liu, Eunwoo Song
While universal vocoders have achieved proficient waveform generation across diverse voices, their integration into text-to-speech (TTS) tasks often results in degraded synthetic q…