activity
20202025
most citedToward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims

219 citations · 446 across the 14 of their papers we have counts for

collaborators
Showing cs.CYShow all

5 papers · 1 filter

cs.CY2024★ 51 cited

The Ethics of Advanced AI Assistants

Iason Gabriel, Arianna Manzini, Geoff Keeling +54

This paper focuses on the opportunities and the ethical and societal risks posed by advanced AI assistants. We define advanced AI assistants as artificial agents with natural langu…

cs.CY2023★ 20 cited

International Institutions for Advanced AI

Lewis Ho, Joslyn Barnhart, Robert Trager +8

International institutions may have an important role to play in ensuring advanced AI systems benefit humanity. International collaborations can unlock AI's ability to further sust…

cs.CY2020★ 219 cited

Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims

Miles Brundage, Shahar Avin, Jasmine Wang +56

With the recent wave of progress in artificial intelligence (AI) has come a growing awareness of the large-scale impacts of AI systems, and recognition that existing regulations an…

cs.CY2020★ 5 cited

The Windfall Clause: Distributing the Benefits of AI for the Common Good

Cullen O'Keefe, Peter Cihon, Ben Garfinkel +3

As the transformative potential of AI has become increasingly salient as a matter of public and political interest, there has been growing discussion about the need to ensure that…

cs.CY2020★ 8 cited

The Offense-Defense Balance of Scientific Knowledge: Does Publishing AI Research Reduce Misuse?

Toby Shevlane, Allan Dafoe

There is growing concern over the potential misuse of artificial intelligence (AI) research. Publishing scientific research can facilitate misuse of the technology, but the researc…