2 papers
cs.CY2024
What AI evaluations for preventing catastrophic risks can and cannot do
Peter Barnett, Lisa Thiergart
AI evaluations are an important component of the AI governance toolkit, underlying current approaches to safety cases for preventing catastrophic risks. Our paper examines what the…
cs.AI2024
Declare and Justify: Explicit assumptions in AI evaluations are necessary for effective regulation
Peter Barnett, Lisa Thiergart
As AI systems advance, AI evaluations are becoming an important pillar of regulations for ensuring safety. We argue that such regulation should require developers to explicitly ide…