most citedRegion-Aware Diffusion for Zero-shot Text-driven Image Editing

7 citations · 22 across the 10 of their papers we have counts for

collaborators

10 papers

cs.CV2024

Interactive Visual Assessment for Text-to-Image Generation Models

Xiaoyue Mi, Fan Tang, Juan Cao +5

Visual generation models have achieved remarkable progress in computer graphics applications but still face significant challenges in real-world deployment. Current assessment appr…

cs.CV2024

HeadRouter: A Training-free Image Editing Framework for MM-DiTs by Adaptively Routing Attention Heads

Yu Xu, Fan Tang, Juan Cao +5

Diffusion Transformers (DiTs) have exhibited robust capabilities in image generation tasks. However, accurate text-guided image editing for multimodal DiTs (MM-DiTs) still poses a…

cs.CV2024

Break-for-Make: Modular Low-Rank Adaptations for Composable Content-Style Customization

Yu Xu, Fan Tang, Juan Cao +5

Personalized generation paradigms empower designers to customize visual intellectual properties with the help of textual descriptions by tuning or adapting pre-trained text-to-imag…

cs.CV2024

Make-Your-Anchor: A Diffusion-based 2D Avatar Generation Framework

Ziyao Huang, Fan Tang, Yong Zhang +4

Despite the remarkable process of talking-head-based avatar-creating solutions, directly generating anchor-style videos with full-body motions remains challenging. In this study, w…

cs.CV20242 cited

Image Collage on Arbitrary Shape via Shape-Aware Slicing and Optimization

Dong-Yi Wu, Thi-Ngoc-Hanh Le, Sheng-Yi Yao +2

Image collage is a very useful tool for visualizing an image collection. Most of the existing methods and commercial applications for generating image collages are designed on simp…

cs.CV20232 cited

Regenerating Arbitrary Video Sequences with Distillation Path-Finding

Thi-Ngoc-Hanh Le, Sheng-Yi Yao, Chun-Te Wu +1

If the video has long been mentioned as a widespread visualization form, the animation sequence in the video is mentioned as storytelling for people. Producing an animation require…