15 citations · 15 across the 1 of their papers we have counts for
1 paper
Rohin Shah, Dmitrii Krasheninnikov, Jordan Alexander +2
Reinforcement learning (RL) agents optimize only the features specified in a reward function and are indifferent to anything left out inadvertently. This means that we must not onl…