1 paper
Qianxun Xu, Chenxi Song, Yujun Cai +1
Recent advances in text-to-video diffusion models have enabled high-fidelity and temporally coherent videos synthesis. However, current models are predominantly optimized for singl…