12 citations · 21 across the 4 of their papers we have counts for
4 papers
Data-driven grapheme-to-phoneme representations for a lexicon-free text-to-speech
Abhinav Garg, Jiyeon Kim, Sushil Khyalia +2
Grapheme-to-Phoneme (G2P) is an essential first step in any modern, high-quality Text-to-Speech (TTS) system. Most of the current G2P systems rely on carefully hand-crafted lexicon…
Time-Varying Quasi-Closed-Phase Analysis for Accurate Formant Tracking in Speech Signals
Dhananjaya Gowda, Sudarsana Reddy Kadiri, Brad Story +1
In this paper, we propose a new method for the accurate estimation and tracking of formants in speech signals using time-varying quasi-closed-phase (TVQCP) analysis. Conventional f…
Refining a Deep Learning-based Formant Tracker using Linear Prediction Methods
Paavo Alku, Sudarsana Reddy Kadiri, Dhananjaya Gowda
In this study, formant tracking is investigated by refining the formants tracked by an existing data-driven tracker, DeepFormants, using the formants estimated in a model-driven ma…
Mitigating the Exposure Bias in Sentence-Level Grapheme-to-Phoneme (G2P) Transduction
Eunseop Yoon, Hee Suk Yoon, Dhananjaya Gowda +7
Text-to-Text Transfer Transformer (T5) has recently been considered for the Grapheme-to-Phoneme (G2P) transduction. As a follow-up, a tokenizer-free byte-level model based on T5 re…