Zuda Yu, Qianhui Xu, Ting Chen +5
Flow Matching (FM) has emerged as a powerful paradigm for speech generation but remains constrained by high inference latency and timbre leakage. To address these bottlenecks, we p…