3 papers
cs.AI2026
Harmonizing AI Safety Thresholds
Wilber Sean Anterola, Matthew Ball, Luis F. Lafuerza +1
Frontier AI companies have published capability thresholds that differ substantially, making it difficult for third parties to verify whether a threshold has been crossed or to com…
cs.CY2025
The AI Risk Spectrum: From Dangerous Capabilities to Existential Threats
Markov Grey, Charbel-Raphaël Segerie
As AI systems become more capable, integrated, and widespread, understanding the associated risks becomes increasingly important. This paper maps the full spectrum of AI risks, fro…
cs.AI2025
Safety by Measurement: A Systematic Literature Review of AI Safety Evaluation Methods
Markov Grey, Charbel-Raphaël Segerie
As frontier AI systems advance toward transformative capabilities, we need a parallel transformation in how we measure and evaluate these systems to ensure safety and inform govern…