activity
20202026
most citedToward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims

219 citations · 337 across the 9 of their papers we have counts for

collaborators
Showing cs.CYShow all

7 papers · 1 filter

cs.CY2026★ 1 cited

Frontier AI Auditing: Toward Rigorous Third-Party Assessment of Safety and Security Practices at Leading AI Companies

Miles Brundage, Noemi Dreksler, Aidan Homewood +45

We outline a vision for frontier AI auditing, which we define as rigorous third-party verification of frontier AI developers' safety and security claims, and evaluation of their sy…

cs.CY2026

Legal Alignment for Safe and Ethical AI

Noam Kolt, Nicholas Caputo, Jack Boeglin +14

Alignment of artificial intelligence (AI) encompasses the normative problem of specifying how AI systems should act and the technical problem of ensuring AI systems comply with tho…

cs.CY2024★ 22 cited

Computing Power and the Governance of Artificial Intelligence

Girish Sastry, Lennart Heim, Haydn Belfield +16

Computing power, or "compute," is crucial for the development and deployment of artificial intelligence (AI) capabilities. As a result, governments and companies have started to le…

cs.CY2023★ 13 cited

Confidence-Building Measures for Artificial Intelligence: Workshop Proceedings

Sarah Shoker, Andrew Reddie, Sarah Barrington +20

Foundation models could eventually introduce several pathways for undermining state security: accidents, inadvertent escalation, unintentional conflict, the proliferation of weapon…

cs.CY2023★ 76 cited

Frontier AI Regulation: Managing Emerging Risks to Public Safety

Markus Anderljung, Joslyn Barnhart, Anton Korinek +21

Advanced AI models hold the promise of tremendous benefits for humanity, but society needs to proactively manage the accompanying risks. In this paper, we focus on what we term "fr…

cs.CY2020★ 219 cited

Toward Trustworthy AI Development: Mechanisms for Supporting Verifiable Claims

Miles Brundage, Shahar Avin, Jasmine Wang +56

With the recent wave of progress in artificial intelligence (AI) has come a growing awareness of the large-scale impacts of AI systems, and recognition that existing regulations an…