3 papers
cs.CV2024
Motion Inversion for Video Customization
Luozhou Wang, Ziyang Mai, Guibao Shen +6
In this work, we present a novel approach for motion customization in video generation, addressing the widespread gap in the exploration of motion representation within video gener…
cs.CV2024
Text-Anchored Score Composition: Tackling Condition Misalignment in Text-to-Image Diffusion Models
Luozhou Wang, Guibao Shen, Wenhang Ge +3
Text-to-image diffusion models have advanced towards more controllable generation via supporting various additional conditions (e.g.,depth map, bounding box) beyond text. However,…
cs.CV2024
SG-Adapter: Enhancing Text-to-Image Generation with Scene Graph Guidance
Guibao Shen, Luozhou Wang, Jiantao Lin +9
Recent advancements in text-to-image generation have been propelled by the development of diffusion models and multi-modality learning. However, since text is typically represented…