117 citations · 183 across the 10 of their papers we have counts for
Showing 2018Show all
2 papers · 1 filter
cs.LG2018
Scalable agent alignment via reward modeling: a research direction
Jan Leike, David Krueger, Tom Everitt +3
One obstacle to applying reinforcement learning algorithms to real-world problems is the lack of suitable reward functions. Designing such reward functions is difficult in part bec…
cs.AI2018
AGI Safety Literature Review
Tom Everitt, Gary Lea, Marcus Hutter
The development of Artificial General Intelligence (AGI) promises to be a major event. Along with its many potential benefits, it also raises serious safety concerns (Bostrom, 2014…