2 papers
cs.CV2025
Canvas-to-Image: Compositional Image Generation with Multimodal Controls
Yusuf Dalva, Guocheng Gordon Qian, Maya Goldenberg +5
While modern diffusion models excel at generating high-quality and diverse images, they still struggle with high-fidelity compositional and multimodal control, particularly when us…
cs.CV2025
LayerComposer: Multi-Human Personalized Generation via Layered Canvas
Guocheng Gordon Qian, Ruihang Zhang, Tsai-Shien Chen +10
Despite their impressive visual fidelity, existing personalized image generators lack interactive control over spatial composition and scale poorly to multiple humans. To address t…