1 paper
Ali Larian, Qian Lin, Chang Zong Wu +1
As autonomous agents are increasingly deployed across diverse operational contexts, aligning their behavior with human intent demands reward functions that remain robust to such ch…