1 paper
Biao Liu, Ning Xu, Junming Yang +2
Large Language Models (LLMs) have achieved remarkable success across diverse natural language tasks, yet the reward models employed for aligning LLMs often encounter challenges of…