14 citations · 30 across the 6 of their papers we have counts for
1 paper · 1 filter
Jiyoung Lee, Joon Son Chung, Soo-Whan Chung
The goal of this work is zero-shot text-to-speech synthesis, with speaking styles and voices learnt from facial characteristics. Inspired by the natural fact that people can imagin…