4 papers
Lessons from External Review of DeepMind's Scheming Inability Safety Case
Stephen Barrett, Francisco Javier Campos Zabala, Sean P. Fillingham +4
Safety cases for frontier AI systems should provide a convincing argument, supported by evidence, that the risk of harm is within an acceptable bound. When developers author their…
The Competence Shadow: Theory and Bounds of AI Assistance in Safety Engineering
Umair Siddique
As AI assistants become integrated into safety engineering workflows for Physical AI systems, a critical question emerges: does AI assistance improve safety analysis quality, or in…
A Digital Twin Framework for Metamorphic Testing of Autonomous Driving Systems Using Generative Model
Tony Zhang, Burak Kantarci, Umair Siddique
Ensuring the safety of self-driving cars remains a major challenge due to the complexity and unpredictability of real-world driving environments. Traditional testing methods face s…
AI-Augmented Metamorphic Testing for Comprehensive Validation of Autonomous Vehicles
Tony Zhang, Burak Kantarci, Umair Siddique
Self-driving cars have the potential to revolutionize transportation, but ensuring their safety remains a significant challenge. These systems must navigate a variety of unexpected…