2 papers
cs.CY2025
Societal Capacity Assessment Framework: Measuring Resilience to Inform Advanced AI Risk Management
Milan Gandhi, Peter Cihon, Owen Larter +1
Risk assessments for advanced AI systems require evaluating both the models themselves and their deployment contexts. We introduce the Societal Capacity Assessment Framework (SCAF)…
cs.CY2025
Who Should Run Advanced AI Evaluations -- AISIs?
Merlin Stein, Milan Gandhi, Theresa Kriecherbauer +2
Artificial Intelligence (AI) Safety Institutes and governments worldwide are deciding whether they evaluate advanced AI themselves, support a private evaluation ecosystem or do bot…