Showing cs.CYShow all
2 papers · 1 filter
cs.CY2026
Lessons from External Review of DeepMind's Scheming Inability Safety Case
Stephen Barrett, Francisco Javier Campos Zabala, Sean P. Fillingham +4
Safety cases for frontier AI systems should provide a convincing argument, supported by evidence, that the risk of harm is within an acceptable bound. When developers author their…
cs.CY2025
Where AI Assurance Might Go Wrong: Initial lessons from engineering of critical systems
Robin Bloomfield, John Rushby
We draw on our experience working on system and software assurance and evaluation for systems important to society to summarise how safety engineering is performed in traditional c…