1 paper · 1 filter
Roger Creus Castanyer, Faisal Mohamed, Pablo Samuel Castro +2
Reinforcement learning (RL) algorithms are highly sensitive to reward function specification, which remains a central challenge limiting their broad applicability. We present ARM-F…