1 paper
Shihao Zhang, Yunzhi Li, Yuguang Yan +4
Recent text-to-video (T2V) diffusion models rely heavily on auxiliary reward signals (e.g., via reward models or DPO) to align generated content with human aesthetics and improve r…