1 paper
Yoshifumi Nakano, Takaaki Saeki, Shinnosuke Takamichi +2
This paper proposes visual-text to speech (vTTS), a method for synthesizing speech from visual text (i.e., text as an image). Conventional TTS converts phonemes or characters into…