8 citations · 8 across the 1 of their papers we have counts for
1 paper
Xinran Liang, Katherine Shu, Kimin Lee +1
Conveying complex objectives to reinforcement learning (RL) agents often requires meticulous reward engineering. Preference-based RL methods are able to learn a more flexible rewar…