2 papers
eess.AS2022
Polyphone disambiguation and accent prediction using pre-trained language models in Japanese TTS front-end
Rem Hida, Masaki Hamada, Chie Kamada +3
Although end-to-end text-to-speech (TTS) models can generate natural speech, challenges still remain when it comes to estimating sentence-level phonetic and prosodic information fr…
cs.CL2018
Dynamic and Static Topic Model for Analyzing Time-Series Document Collections
Rem Hida, Naoya Takeishi, Takehisa Yairi +1
For extracting meaningful topics from texts, their structures should be considered properly. In this paper, we aim to analyze structured time-series documents such as a collection…