3 papers
cs.CV2026
FreeSpec: Training-Free Long Video Generation via Singular-Spectrum Reconstruction
Fangda Chen, Shanshan Zhao, Longrong Yang +3
Video diffusion models perform well in short-video synthesis, but their training-free extension to long videos often suffers from content drift, temporal inconsistency, and over-sm…
cs.CV2024
Ctrl-V: Higher Fidelity Video Generation with Bounding-Box Controlled Object Motion
Ge Ya Luo, Zhi Hao Luo, Anthony Gosselin +2
Controllable video generation has attracted significant attention, largely due to advances in video diffusion models. In domains such as autonomous driving, it is essential to deve…
cs.CV2024
Beyond FVD: Enhanced Evaluation Metrics for Video Generation Quality
Ge Ya Luo, Gian Mario Favero, Zhi Hao Luo +2
The Fréchet Video Distance (FVD) is a widely adopted metric for evaluating video generation distribution quality. However, its effectiveness relies on critical assumptions. Our an…