From the 1 of 4 linked papers with an AI index.
4 papers
Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions
Xin Jin, Huanqia Cai, Zhen Li +9
The paper introduces Z-Reward, a teacher‑student framework that learns to predict full rubric‑aligned score distributions for text‑to‑image generation instead of single scalar rewa…
TBStar-Edit: From Image Editing Pattern Shifting to Consistency Enhancement
Hao Fang, Zechao Zhan, Weixin Feng +3
Recent advances in image generation and editing technologies have enabled state-of-the-art models to achieve impressive results in general domains. However, when applied to e-comme…
MADiff: Text-Guided Fashion Image Editing with Mask Prediction and Attention-Enhanced Diffusion
Zechao Zhan, Dehong Gao, Jinxia Zhang +3
Text-guided image editing model has achieved great success in general domain. However, directly applying these models to the fashion domain may encounter two issues: (1) Inaccurate…
FashionFAE: Fine-grained Attributes Enhanced Fashion Vision-Language Pre-training
Jiale Huang, Dehong Gao, Jinxia Zhang +3
Large-scale Vision-Language Pre-training (VLP) has demonstrated remarkable success in the general domain. However, in the fashion domain, items are distinguished by fine-grained at…