1 paper · 1 filter
Julia Santaniello, Matthew Russell, Benson Jiang +3
Reinforcement Learning from Human Feedback (RLHF) is a methodology that aligns agent behavior with human preferences by integrating user feedback into the agent's training process.…