1 paper
Sara Rajaram, R. James Cotton, Fabian H. Sinz
Preference-based Reinforcement Learning (PbRL) entails a variety of approaches for aligning models with human intent to alleviate the burden of reward engineering. However, most pr…