1 paper · 1 filter
Raghav Sharma, Manan Mehta, Sai Tiger Raina
Reinforcement Learning from Human Feedback (RLHF) is the standard for aligning Large Language Models (LLMs), yet recent progress has moved beyond canonical text-based methods. This…