Fairness Through Counterfactual Utilities
arXiv:2108.05315 · doi:10.1613/jair.1.14238
Abstract
Group fairness definitions such as Demographic Parity and Equal Opportunity make assumptions about the underlying decision-problem that restrict them to classification problems. Prior work has translated these definitions to other machine learning environments, such as unsupervised learning and reinforcement learning, by implementing their closest mathematical equivalent. As a result, there are numerous bespoke interpretations of these definitions. Instead, we provide a generalized set of group fairness definitions that unambiguously extend to all machine learning environments while still retaining their original fairness notions. We derive two fairness principles that enable such a generalized framework. First, our framework measures outcomes in terms of utilities, rather than predictions, and does so for both the decision-algorithm and the individual. Second, our framework considers counterfactual outcomes, rather than just observed outcomes, thus preventing loopholes where fairness criteria are satisfied through self-fulfilling prophecies. We provide concrete examples of how our counterfactual utility fairness framework resolves known fairness issues in classification, clustering, and reinforcement learning problems. We also show that many of the bespoke interpretations of Demographic Parity and Equal Opportunity fit nicely as special cases of our framework.
References in corpus (9)
- Equality of Opportunity in Supervised Learning
- Fairness Testing: Testing Software for Discrimination
- Avoiding Discrimination through Causal Reasoning
- Minimax Pareto Fairness: A Multi Objective Perspective
- Policy Learning for Fairness in Ranking
- Survey on Causal-based Machine Learning Fairness Notions
- Outside the Echo Chamber: Optimizing the Performative Risk
- Trade-offs between Group Fairness Metrics in Societal Resource Allocation
- Fair for All: Best-effort Fairness Guarantees for Classification