2 citations · 7 across the 13 of their papers we have counts for
3 papers · 1 filter
Preliminary suggestions for rigorous GPAI model evaluations
Patricia Paskov, Michael J. Byun, Kevin Wei +1
This document presents a preliminary compilation of general-purpose AI (GPAI) evaluation practices that may promote internal validity, external validity and reproducibility. It inc…
Recommendations and Reporting Checklist for Rigorous & Transparent Human Baselines in Model Evaluations
Kevin L. Wei, Patricia Paskov, Sunishchal Dev +6
In this position paper, we argue that human baselines in foundation model evaluations must be more rigorous and more transparent to enable meaningful comparisons of human vs. AI pe…
In Which Areas of Technical AI Safety Could Geopolitical Rivals Cooperate?
Ben Bucknall, Saad Siddiqui, Lara Thurnherr +19
International cooperation is common in AI research, including between geopolitical rivals. While many experts advocate for greater international cooperation on AI safety to address…