1 paper · 1 filter
Yachen Kang, Diyuan Shi, Jinxin Liu +2
This study focuses on the topic of offline preference-based reinforcement learning (PbRL), a variant of conventional reinforcement learning that dispenses with the need for online…