collaborators

6 papers

cs.CY2026

FLARE-AI: Flaw Reporting for AI

Shayne Longpre, Elaine Zhu, Carson Ezell +15

Flaw reporting for deployed AI systems is fundamental to identifying system failures and improving AI safety. Yet the AI reporting ecosystem is fragmented: researchers who identify…

cs.CY2026

RCTs for Frontier AI Governance: Methodological Challenges and Solutions for Human Uplift Studies

Patricia Paskov, Kevin Wei, Shen Zhou Hong +7

Human uplift studies, or studies that measure the effects of AI access on human performance via randomized controlled trials (RCT) or similar methodologies, increasingly inform fro…

cs.CY2025

Dark Speculation: Combining Qualitative and Quantitative Understanding in Frontier AI Risk Analysis

Daniel Carpenter, Carson Ezell, Pratyush Mallick +1

Estimating catastrophic harms from frontier AI is hindered by deep ambiguity: many of its risks are not only unobserved but unanticipated by analysts. The central limitation of cur…

cs.CY2025

Incident Analysis for AI Agents

Carson Ezell, Xavier Roberts-Gaal, Alan Chan

As AI agents become more widely deployed, we are likely to see an increasing number of incidents: events involving AI agent use that directly or indirectly cause harm. For example,…

cs.MA2025

Multi-Agent Risks from Advanced AI

Lewis Hammond, Alan Chan, Jesse Clifton +41

The rapid development of advanced AI agents and the imminent deployment of many instances of these agents will give rise to multi-agent systems of unprecedented complexity. These s…

cs.SE2025

The AI Agent Index

Stephen Casper, Luke Bailey, Rosco Hunter +12

Leading AI developers and startups are increasingly deploying agentic AI systems that can plan and execute complex tasks with limited human involvement. However, there is currently…