1 paper
Hyomin Kim, Junghye Kim, Joanie Hayoun Chung +4
Reward models for text-to-video (T2V) generation guide post-training but often fail at fine-grained semantic alignment. We trace this to two structural weaknesses in existing reaso…