Showing cs.CVShow all
3 papers · 1 filter
cs.CV2025
Encapsulated Composition of Text-to-Image and Text-to-Video Models for High-Quality Video Synthesis
Tongtong Su, Chengyu Wang, Bingyan Liu +2
In recent years, large text-to-video (T2V) synthesis models have garnered considerable attention for their abilities to generate videos from textual descriptions. However, achievin…
cs.CV2025
Zero-to-Hero: Zero-Shot Initialization Empowering Reference-Based Video Appearance Editing
Tongtong Su, Chengyu Wang, Jun Huang +1
Appearance editing according to user needs is a pivotal task in video editing. Existing text-guided methods often lead to ambiguities regarding user intentions and restrict fine-gr…
cs.CV2025
Understanding Attention Mechanism in Video Diffusion Models
Bingyan Liu, Chengyu Wang, Tongtong Su +4
Text-to-video (T2V) synthesis models, such as OpenAI's Sora, have garnered significant attention due to their ability to generate high-quality videos from a text prompt. In diffusi…