1 paper
Mingi Kwon, Joonghyuk Shin, Jaeseok Jung +2
The intrinsic link between facial motion and speech is often overlooked in generative modeling, where talking head synthesis and text-to-speech (TTS) are typically addressed as sep…