117 citations · 263 across the 17 of their papers we have counts for
1 paper · 1 filter
Victoria Krakovna, Laurent Orseau, Ramana Kumar +2
How can we design safe reinforcement learning agents that avoid unnecessary disruptions to their environment? We show that current approaches to penalizing side effects can introdu…