2 papers
cs.AI2025
Recommendations and Reporting Checklist for Rigorous & Transparent Human Baselines in Model Evaluations
Kevin L. Wei, Patricia Paskov, Sunishchal Dev +6
In this position paper, we argue that human baselines in foundation model evaluations must be more rigorous and more transparent to enable meaningful comparisons of human vs. AI pe…
cs.CY2024
How Do AI Companies "Fine-Tune" Policy? Examining Regulatory Capture in AI Governance
Kevin Wei, Carson Ezell, Nick Gabrieli +1
Industry actors in the United States have gained extensive influence in conversations about the regulation of general-purpose artificial intelligence (AI) systems. Although industr…