1 paper · 1 filter
Chenglong Wang, Yifu Huo, Yang Gan +10
Previous methods evaluate reward models by testing them on a fixed pairwise ranking test set, but they typically do not provide performance information on each preference dimension…