3 papers
cs.CY2026
Prioritization of Risks from Artificial Intelligence: A Delphi Study of 272 International Experts
Alexander K. Saeri, Jess Graham, Michael Noetel +185
Artificial intelligence poses many risks, ranging from familiar present-day harms to unprecedented and potentially catastrophic ones. Effective risk management requires prioritizat…
cs.CY2026
Preregistration for Experiments with AI Agents
Michelle Vaccaro
The proliferation of large language models (LLMs) and autonomous AI agents has given rise to a rapidly growing methodological paradigm: "in silico" behavioral experiments. Original…
cs.CY2026
Evaluating Human-AI Safety: A Framework for Measuring Harmful Capability Uplift
Michelle Vaccaro, Jaeyoon Song, Abdullah Almaatouq +1
Current frontier AI safety evaluations emphasize static benchmarks, third-party annotations, and red-teaming. In this position paper, we argue that AI safety research should focus…