112 citations · 121 across the 2 of their papers we have counts for
2 papers
cs.LG2019★ 9 cited
Generalizing from a few environments in safety-critical reinforcement learning
Zachary Kenton, Angelos Filos, Owain Evans +1
Before deploying autonomous agents in the real world, we need to be confident they will perform safely in novel situations. Ideally, we would expose agents to a very wide range of…
cs.AI2017★ 112 cited
Trial without Error: Towards Safe Reinforcement Learning via Human Intervention
William Saunders, Girish Sastry, Andreas Stuhlmueller +1
AI systems are increasingly applied to complex tasks that involve interaction with humans. During training, such systems are potentially dangerous, as they haven't yet learned to a…