1 paper
Yuxuan Li, Harshith Reddy Kethireddy, Srijita Das
Learning from Preferences in Reinforcement Learning (PbRL) has gained attention recently, as it serves as a natural fit for complicated tasks where the reward function is not easil…