219 citations · 408 across the 4 of their papers we have counts for
7 papers
An Economy of AI Agents
Gillian K. Hadfield, Andrew Koh
In the coming decade, artificially intelligent agents with the ability to plan and execute complex tasks over long time horizons with little direct oversight from humans may be dep…
Gathering Strength, Gathering Storms: The One Hundred Year Study on Artificial Intelligence (AI100) 2021 Study Panel Report
Michael L. Littman, Ifeoma Ajunwa, Guy Berger +14
In September 2021, the "One Hundred Year Study on Artificial Intelligence" project (AI100) issued the second report of its planned long-term periodic assessment of artificial intel…
Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims
Miles Brundage, Shahar Avin, Jasmine Wang +56
With the recent wave of progress in artificial intelligence (AI) has come a growing awareness of the large-scale impacts of AI systems, and recognition that existing regulations an…
Silly rules improve the capacity of agents to learn stable enforcement and compliance behaviors
Raphael Köster, Dylan Hadfield-Menell, Gillian K. Hadfield +1
How can societies learn to enforce and comply with social norms? Here we investigate the learning dynamics and emergence of compliance and enforcement of social norms in a foraging…
The Role of Cooperation in Responsible AI Development
Amanda Askell, Miles Brundage, Gillian Hadfield
In this paper, we argue that competitive pressures could incentivize AI companies to underinvest in ensuring their systems are safe, secure, and have a positive social impact. Ensu…
Legible Normativity for AI Alignment: The Value of Silly Rules
Dylan Hadfield-Menell, McKane Andrus, Gillian K. Hadfield
It has become commonplace to assert that autonomous agents will have to be built to follow human rules of behavior--social norms and laws. But human laws and norms are complex and…