1 citations · 2 across the 8 of their papers we have counts for
Showing cs.SDShow all
2 papers · 1 filter
cs.SD2026
MeloBottleneck: Self-Supervised Melody Skeleton Extraction with a Latent Subsequence Bottleneck
Fan Bu, Rongfeng Li, Linfeng Fan
Melody skeleton extraction aims to derive a shorter melody that preserves structural notes while removing ornaments. Prior methods rely on hand-crafted reduction rules or note-wise…
cs.SD2024
CM-TTS: Enhancing Real Time Text-to-Speech Synthesis Efficiency through Weighted Samplers and Consistency Models
Xiang Li, Fan Bu, Ambuj Mehrish +4
Neural Text-to-Speech (TTS) systems find broad applications in voice assistants, e-learning, and audiobook creation. The pursuit of modern models, like Diffusion Models (DMs), hold…