23 citations · 25 across the 7 of their papers we have counts for
1 paper · 1 filter
Hanfang Lyu, Yuanchen Bai, Xin Liang +7
Preference-based learning aims to align robot task objectives with human values. One of the most common methods to infer human preferences is by pairwise comparisons of robot task…