118 citations · 213 across the 3 of their papers we have counts for
1 paper · 1 filter
Shang Liu, Yu Pan, Guanting Chen +1
Learning a reward model (RM) from human preferences has been an important component in aligning large language models (LLMs). The canonical setup of learning RMs from pairwise pref…