4.3k citations · 6.1k across the 19 of their papers we have counts for
Showing cs.CYShow all
2 papers · 1 filter
cs.CY2021★ 71 cited
Institutionalising Ethics in AI through Broader Impact Requirements
Carina Prunkl, Carolyn Ashurst, Markus Anderljung +3
Turning principles into practice is one of the most pressing challenges of artificial intelligence (AI) governance. In this article, we reflect on a novel governance initiative by…
cs.CY2019
Learning Human Objectives by Evaluating Hypothetical Behavior
Siddharth Reddy, Anca D. Dragan, Sergey Levine +2
We seek to align agent behavior with a user's objectives in a reinforcement learning setting with unknown dynamics, an unknown reward function, and unknown unsafe states. The user…