1 paper
Yijun Liu, Jie Huang, Zeyue Xue +5
Reward models guide text-to-image (T2I) systems toward outputs aligned with human preferences. However, typical reward models such as HPSv3 are trained on pre-annotated data from e…