1 paper
Yi-Shiuan Tung, Yuni Wu, Wei Jiang +2
Learning reward functions from human preferences is a widely used approach for aligning robot behavior with user expectations in human-robot interaction. Most existing approaches a…