4 citations · 4 across the 5 of their papers we have counts for
5 papers · 1 filter
FlowBlending: Stage-Aware Multi-Model Sampling for Fast and High-Fidelity Video Generation
Jibin Song, Mingi Kwon, Jaeseok Jeong +1
In this work, we show that the impact of model capacity varies across timesteps: it is crucial for the early and late stages but largely negligible during the intermediate stage. A…
Balanced conic rectified flow
Shin Seong Kim, Mingi Kwon, Jaeseok Jeong +1
Rectified flow is a generative model that learns smooth transport mappings between two distributions through an ordinary differential equation (ODE). Unlike diffusion-based generat…
StyleKeeper: Prevent Content Leakage using Negative Visual Query Guidance
Jaeseok Jeong, Junho Kim, Gayoung Lee +2
In the domain of text-to-image generation, diffusion models have emerged as powerful tools. Recently, studies on visual prompting, where images are used as prompts, have enabled mo…
Syncphony: Synchronized Audio-to-Video Generation with Diffusion Transformers
Jibin Song, Mingi Kwon, Jaeseok Jeong +1
Text-to-video and image-to-video generation have made rapid progress in visual quality, but they remain limited in controlling the precise timing of motion. In contrast, audio prov…
Visual Style Prompting with Swapping Self-Attention
Jaeseok Jeong, Junho Kim, Yunjey Choi +2
In the evolving domain of text-to-image generation, diffusion models have emerged as powerful tools in content creation. Despite their remarkable capability, existing models still…