1 paper
Haoran Qin, Renlong Wu, Tianyu Huang +3
While diffusion-based text-to-video (T2V) models have demonstrated impressive capability in generating realistic and temporally coherent videos, they often fail to respect fundamen…