1 paper
Se Jin Park, Minsu Kim, Joanna Hong +2
The challenge of talking face generation from speech lies in aligning two different modal information, audio and video, such that the mouth region corresponds to input audio. Previ…