203 citations · 580 across the 30 of their papers we have counts for
12 papers · 1 filter
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
Yusuf Dalva, Guocheng Gordon Qian, Maya Goldenberg +5
While modern diffusion models excel at generating high-quality and diverse images, they still struggle with high-fidelity compositional and multimodal control, particularly when us…
Preventing Shortcuts in Adapter Training via Providing the Shortcuts
Anujraaj Argo Goyal, Guocheng Gordon Qian, Huseyin Coskun +8
Adapter-based training has emerged as a key mechanism for extending the capabilities of powerful foundation image generators, enabling personalized and stylized text-to-image synth…
ComposeMe: Attribute-Specific Image Prompts for Controllable Human Image Generation
Guocheng Gordon Qian, Daniil Ostashev, Egor Nemchinov +4
Generating high-fidelity images of humans with fine-grained control over attributes such as hairstyle and clothing remains a core challenge in personalized text-to-image synthesis.…
Scaling Group Inference for Diverse and High-Quality Generation
Gaurav Parmar, Or Patashnik, Daniil Ostashev +4
Generative models typically sample outputs independently, and recent inference-time guidance and scaling algorithms focus on improving the quality of individual samples. However, i…
Zero-Shot Dynamic Concept Personalization with Grid-Based LoRA
Rameen Abdal, Or Patashnik, Ekaterina Deyneka +5
Recent advances in text-to-video generation have enabled high-quality synthesis from text and image prompts. While the personalization of dynamic concepts, which capture subject-sp…
3D PixBrush: Image-Guided Local Texture Synthesis
Dale Decatur, Itai Lang, Kfir Aberman +1
We present 3D PixBrush, a method for performing image-driven edits of local regions on 3D meshes. 3D PixBrush predicts a localization mask and a synthesized texture that faithfully…