Showing cs.AIShow all
2 papers · 1 filter
cs.AI2026
Harmonizing AI Safety Thresholds
Wilber Sean Anterola, Matthew Ball, Luis F. Lafuerza +1
Frontier AI companies have published capability thresholds that differ substantially, making it difficult for third parties to verify whether a threshold has been crossed or to com…
cs.AI2025
Safety by Measurement: A Systematic Literature Review of AI Safety Evaluation Methods
Markov Grey, Charbel-Raphaël Segerie
As frontier AI systems advance toward transformative capabilities, we need a parallel transformation in how we measure and evaluate these systems to ensure safety and inform govern…