7 citations · 22 across the 10 of their papers we have counts for
10 papers
Interactive Visual Assessment for Text-to-Image Generation Models
Xiaoyue Mi, Fan Tang, Juan Cao +5
Visual generation models have achieved remarkable progress in computer graphics applications but still face significant challenges in real-world deployment. Current assessment appr…
HeadRouter: A Training-free Image Editing Framework for MM-DiTs by Adaptively Routing Attention Heads
Yu Xu, Fan Tang, Juan Cao +5
Diffusion Transformers (DiTs) have exhibited robust capabilities in image generation tasks. However, accurate text-guided image editing for multimodal DiTs (MM-DiTs) still poses a…
Break-for-Make: Modular Low-Rank Adaptations for Composable Content-Style Customization
Yu Xu, Fan Tang, Juan Cao +5
Personalized generation paradigms empower designers to customize visual intellectual properties with the help of textual descriptions by tuning or adapting pre-trained text-to-imag…
Make-Your-Anchor: A Diffusion-based 2D Avatar Generation Framework
Ziyao Huang, Fan Tang, Yong Zhang +4
Despite the remarkable process of talking-head-based avatar-creating solutions, directly generating anchor-style videos with full-body motions remains challenging. In this study, w…
Image Collage on Arbitrary Shape via Shape-Aware Slicing and Optimization
Dong-Yi Wu, Thi-Ngoc-Hanh Le, Sheng-Yi Yao +2
Image collage is a very useful tool for visualizing an image collection. Most of the existing methods and commercial applications for generating image collages are designed on simp…
Regenerating Arbitrary Video Sequences with Distillation Path-Finding
Thi-Ngoc-Hanh Le, Sheng-Yi Yao, Chun-Te Wu +1
If the video has long been mentioned as a widespread visualization form, the animation sequence in the video is mentioned as storytelling for people. Producing an animation require…