1 paper
Benjamin Poole, Minwoo Lee
Reinforcement learning (RL) research has increasingly shifted focus towards alignment, ensuring agents learn behaviors adhering to human values. While human demonstrations and feed…