1 paper · 1 filter
Nan Lu, Ethan Lee, Ethan X. Fang +1
Reinforcement Learning from Human Feedback (RLHF) has become a pivotal paradigm in artificial intelligence to align large models with human preferences. In this paper, we propose a…