1 paper
Yunji Chu, Yunseob Shim, Unsang Park
We propose FEIM-TTS, an innovative zero-shot text-to-speech (TTS) model that synthesizes emotionally expressive speech, aligned with facial images and modulated by emotion intensit…