3 papers
eess.AS2024
SSDM: Scalable Speech Dysfluency Modeling
Jiachen Lian, Xuanru Zhou, Zoe Ezzes +6
Speech dysfluency modeling is the core module for spoken language learning, and speech therapy. However, there are three challenges. First, current state-of-the-art solutions\cite{…
cs.SD2024
VoxGenesis: Unsupervised Discovery of Latent Speaker Manifold for Speech Synthesis
Weiwei Lin, Chenhang He, Man-Wai Mak +2
Achieving nuanced and accurate emulation of human voice has been a longstanding goal in artificial intelligence. Although significant progress has been made in recent years, the ma…
cs.CL2024
Towards Hierarchical Spoken Language Dysfluency Modeling
Jiachen Lian, Gopala Anumanchipalli
Speech disfluency modeling is the bottleneck for both speech therapy and language learning. However, there is no effective AI solution to systematically tackle this problem. We sol…