7 citations · 19 across the 11 of their papers we have counts for
Showing eess.ASShow all
3 papers · 1 filter
eess.AS2025
Speaker-Conditioned Phrase Break Prediction for Text-to-Speech with Phoneme-Level Pre-trained Language Model
Dong Yang, Yuki Saito, Takaaki Saeki +4
This paper advances phrase break prediction (also known as phrasing) in multi-speaker text-to-speech (TTS) systems. We integrate speaker-specific features by leveraging speaker emb…
eess.AS2023
Duration-aware pause insertion using pre-trained language model for multi-speaker text-to-speech
Dong Yang, Tomoki Koriyama, Yuki Saito +3
Pause insertion, also known as phrase break prediction and phrasing, is an essential part of TTS systems because proper pauses with natural duration significantly enhance the rhyth…
eess.AS2022★ 1 cited
Low-Complexity Acoustic Echo Cancellation with Neural Kalman Filtering
Dong Yang, Fei Jiang, Wei Wu +2
The Kalman filter has been adopted in acoustic echo cancellation due to its robustness to double-talk, fast convergence, and good steady-state performance. The performance of Kalma…