From the 1 of 7 linked papers with an AI index.
1 paper · 1 filter
Qi Zhou, Jie Zhang, Dongxia Wang +5
Human preference plays a crucial role in the refinement of large language models (LLMs). However, collecting human preference feedback is costly and most existing datasets neglect…