2 papers
cs.LG2026
Evaluation Metrics as Averaged Outcomes of Fair Gambles
Rabanus Derr, Robert C. Williamson
In the current practices of machine learning, the evaluation of forecasts has become a cornerstone of scientific progress. A multitude of evaluation metrics have been suggested and…
cs.LG2025
Three Types of Calibration with Properties and their Semantic and Formal Relationships
Rabanus Derr, Jessie Finocchiaro, Robert C. Williamson
Fueled by discussions around "trustworthiness" and algorithmic fairness, calibration of predictive systems has regained scholars attention. The vanilla definition and understanding…