works on

From the 1 of 12 linked papers with an AI index.

activity
20242026
collaborators

12 papers

cs.CY2026

Underwriting the Agent Economy: The Blueprint for an AI Insurance Stack

Cristian Trout, Sanmi Koyejo, Sasha Romanosky +34

The paper proposes a comprehensive AI insurance framework to price and manage risks from the emerging AI agent economy, outlining an eight‑component stack for data collection, mode…

cs.CY2026

RCTs for Frontier AI Governance: Methodological Challenges and Solutions for Human Uplift Studies

Patricia Paskov, Kevin Wei, Shen Zhou Hong +7

Human uplift studies, or studies that measure the effects of AI access on human performance via randomized controlled trials (RCT) or similar methodologies, increasingly inform fro…

cs.CY2026

The 2025 AI Agent Index: Documenting Technical and Safety Features of Deployed Agentic AI Systems

Leon Staufer, Kevin Feng, Kevin Wei +6

Agentic AI systems are increasingly capable of performing professional and personal tasks with limited human involvement. However, tracking these developments is difficult because…

cs.CY2026

Designing Incident Reporting Systems for Harms from General-Purpose AI

Kevin Wei, Lennart Heim

We introduce a conceptual framework and provide considerations for the institutional design of AI incident reporting systems, i.e., processes for collecting information about safet…

cs.LG2026

From Human-Level AI Tales to AI Leveling Human Scales

Peter Romero, Fernando Martínez-Plumed, Zachary R. Tidler +11

Comparing AI models to "human level" is often misleading when benchmark scores are incommensurate or human baselines are drawn from a narrow population. To address this, we propose…

cs.CY2025

Preliminary suggestions for rigorous GPAI model evaluations

Patricia Paskov, Michael J. Byun, Kevin Wei +1

This document presents a preliminary compilation of general-purpose AI (GPAI) evaluation practices that may promote internal validity, external validity and reproducibility. It inc…