1 paper · 1 filter
Siddhant Bhambri, Mudit Verma, Upasana Biswas +2
Preference-based Reinforcement Learning (PbRL) has made significant strides in single-agent settings, but has not been studied for multi-agent frameworks. On the other hand, modeli…