1 citations · 4 across the 11 of their papers we have counts for
1 paper · 1 filter
Jin Peng Zhou, Katie Z Luo, Jingwen Gu +3
This paper presents a novel approach to aligning large language models (LLMs) with individual human preferences, sometimes referred to as Reinforcement Learning from \textit{Person…