1 paper
Jin Peng Zhou, Katie Z Luo, Jingwen Gu +3
This paper presents a novel approach to aligning large language models (LLMs) with individual human preferences, sometimes referred to as Reinforcement Learning from \textit{Person…