4 papers
GMO-EDIT: Grounded Multi-Operation Editing for E-Commerce Images
Zipeng Guo, Xiaoan Liu, Lichen Ma +9
Real-world e-commerce image editing often requires multiple, localized, and auditable operations rather than global restyling. This compositional nature poses a dual challenge: mod…
HyperDiT: Hyper-Connected Transformers for High-Fidelity Pixel-Space Diffusion
Yu He, Lichen Ma, Zipeng Guo +5
Pixel-space diffusion models bypass the reconstruction bottleneck of Variational Autoencoders (VAEs) but face a fundamental "granularity dilemma": capturing global semantics favors…
LiWi: Layering in the Wild
Yu He, Fang Li, Haoyang Tong +7
Recent advances in generative models have empowered impressive layered image generation, yet their success is largely confined to graphic design domains. The layering of in-the-wil…
Fashion130K: An E-commerce Fashion Dataset for Outfit Generation with Unified Multi-modal Condition
Yu He, Ting Zhu, Yichun Liu +6
Recent research work on fashion outfit generation focuses on promoting visual consistency of garments by leveraging key information from reference image and text prompt. However, t…