1 paper
Fangda Chen, Shanshan Zhao, Longrong Yang +3
Video diffusion models perform well in short-video synthesis, but their training-free extension to long videos often suffers from content drift, temporal inconsistency, and over-sm…