reasoning internalization 1reward modeling 1score distribution 1teacher-student framework 1text-to-image generation 1
From the 1 of 3 linked papers with an AI index.
3 papers
cs.CV2026
Z-Reward: Beyond Scalar Rewards by Internalizing Reasoning into Score Distributions
Xin Jin, Huanqia Cai, Zhen Li +9
The paper introduces Z-Reward, a teacher‑student framework that learns to predict full rubric‑aligned score distributions for text‑to‑image generation instead of single scalar rewa…
cs.CV2026
High-Fidelity Two-Step Image Generation via Teacher-Aligned End-to-End Distillation
Dongyang Liu, Ruoyi Du, David Liu +7
Few-step diffusion distillation has become increasingly mature for 4-8-step generation, yet pushing further to 2 steps remains challenging. In this work, we introduce Z-Image Turbo…
cs.CV2026
D-OPSD: On-Policy Self-Distillation for Continuously Tuning Step-Distilled Diffusion Models
Dengyang Jiang, Xin Jin, Dongyang Liu +9
The landscape of high-performance image generation models is currently shifting from the inefficient multi-step ones to the efficient few-step counterparts (e.g, Z-Image-Turbo and…