1 paper
Minu Kim, Yongsik Lee, Sehyeok Kang +3
We present Preference Flow Matching (PFM), a new framework for preference-based reinforcement learning (PbRL) that streamlines the integration of preferences into an arbitrary clas…