2 papers
cs.CY2026
Principles and Guidelines for Randomized Controlled Trials in AI Evaluation
Christopher Kelly, Angelica Chowdhury, Alexandra Campili +5
This work establishes a framework for standardizing AI evaluation RCTs (sometimes called human uplift studies). Drawing on established practices from disciplines with established R…
cs.HC2026
How people use Copilot for Health
Beatriz Costa-Gomes, Pavel Tolmachev, Eloise Taysom +15
We analyze over 500,000 de-identified health-related conversations with Microsoft Copilot from January 2026 to characterize what people ask conversational AI about health. We devel…