1 paper
Haiyang Liu, Xingchao Yang, Tomoya Akiyama +4
We present TANGO, a framework for generating co-speech body-gesture videos. Given a few-minute, single-speaker reference video and target speech audio, TANGO produces high-fidelity…