1 paper
Jingyi Huang, Ruohan Zong, Yujun Feng +3
Reinforcement Learning from Human Feedback (RLHF) is critical for aligning Large Language Models (LLMs) with human preferences. However, its efficacy is often compromised by the in…