1 paper · 1 filter
Benjamin Poole, Minwoo Lee
Reinforcement learning (RL) research has increasingly shifted focus towards alignment, ensuring agents learn behaviors adhering to human values. While human demonstrations and feed…