1 citations · 1 across the 1 of their papers we have counts for
1 paper · 1 filter
Hao Lang, Fei Huang, Yongbin Li
Reinforcement learning from human feedback (RLHF) has emerged as an effective approach to aligning large language models (LLMs) to human preferences. RLHF contains three steps, i.e…