1 paper
Rongxiang Zhang, Songhua Liu
Long-form and real-time talking-head generation remains challenging due to a latency-quality trade-off: inefficient multi-step diffusion prohibits streaming generation, whereas rea…