3 papers
cs.AI2026
AI Evaluation Should Work With Humans
Jan Kulveit, Gavin Leech, Tomáš Gavenčiak +1
This position paper argues that the dominant paradigm of AI evaluation (which focuses on superhuman autonomous performance and so implicitly targets the goal of replacing humans) i…
cs.LG2020
Legally grounded fairness objectives
Dylan Holden-Sim, Gavin Leech, Laurence Aitchison
Recent work has identified a number of formally incompatible operational measures for the unfairness of a machine learning (ML) system. As these measures all capture intuitively de…
stat.AP2020
How Robust are the Estimated Effects of Nonpharmaceutical Interventions against COVID-19?
Mrinank Sharma, Sören Mindermann, Jan Markus Brauner +7
To what extent are effectiveness estimates of nonpharmaceutical interventions (NPIs) against COVID-19 influenced by the assumptions our models make? To answer this question, we inv…