collaborators

5 papers

cs.CY2026

Lessons from External Review of DeepMind's Scheming Inability Safety Case

Stephen Barrett, Francisco Javier Campos Zabala, Sean P. Fillingham +4

Safety cases for frontier AI systems should provide a convincing argument, supported by evidence, that the risk of harm is within an acceptable bound. When developers author their…

cs.CY2026

STAMP/STPA Informed Characterization of Factors Leading to Loss of Control in AI Systems

Steve Barrett, Anna Bruvere, Sean P. Fillingham +2

A major concern amongst AI safety practitioners is the possibility of loss of control, whereby humans lose the ability to exert control over increasingly advanced AI systems. The r…

cs.CY2025

Toward Quantitative Modeling of Cybersecurity Risks Due to AI Misuse

Steve Barrett, Malcolm Murray, Otter Quarks +17

Advanced AI systems offer substantial benefits but also introduce risks. In 2025, AI-enabled cyber offense has emerged as a concrete example. This technical report applies a quanti…

cs.CY2025

A Methodology for Quantitative AI Risk Modeling

Malcolm Murray, Steve Barrett, Henry Papadatos +5

Although general-purpose AI systems offer transformational opportunities in science and industry, they simultaneously raise critical concerns about safety, misuse, and potential lo…

cs.CY2025

The Role of Risk Modeling in Advanced AI Risk Management

Chloé Touzet, Henry Papadatos, Malcolm Murray +6

Rapidly advancing artificial intelligence (AI) systems introduce novel, uncertain, and potentially catastrophic risks. Managing these risks requires a mature risk-management infras…