2 papers
cs.CV2025
FlowBlending: Stage-Aware Multi-Model Sampling for Fast and High-Fidelity Video Generation
Jibin Song, Mingi Kwon, Jaeseok Jeong +1
In this work, we show that the impact of model capacity varies across timesteps: it is crucial for the early and late stages but largely negligible during the intermediate stage. A…
cs.CV2025
Syncphony: Synchronized Audio-to-Video Generation with Diffusion Transformers
Jibin Song, Mingi Kwon, Jaeseok Jeong +1
Text-to-video and image-to-video generation have made rapid progress in visual quality, but they remain limited in controlling the precise timing of motion. In contrast, audio prov…