The Fair Game: Auditing & Debiasing AI Algorithms Over Time
arXiv:2508.06443 · doi:10.1017/cfl.2025.8
Abstract
An emerging field of AI, namely Fair Machine Learning (ML), aims to quantify different types of bias (also known as unfairness) exhibited in the predictions of ML algorithms, and to design new algorithms to mitigate them. Often, the definitions of bias used in the literature are observational, i.e. they use the input and output of a pre-trained algorithm to quantify a bias under concern. In reality,these definitions are often conflicting in nature and can only be deployed if either the ground truth is known or only in retrospect after deploying the algorithm. Thus,there is a gap between what we want Fair ML to achieve and what it does in a dynamic social environment. Hence, we propose an alternative dynamic mechanism,"Fair Game",to assure fairness in the predictions of an ML algorithm and to adapt its predictions as the society interacts with the algorithm over time. "Fair Game" puts together an Auditor and a Debiasing algorithm in a loop around an ML algorithm. The "Fair Game" puts these two components in a loop by leveraging Reinforcement Learning (RL). RL algorithms interact with an environment to take decisions, which yields new observations (also known as data/feedback) from the environment and in turn, adapts future decisions. RL is already used in algorithms with pre-fixed long-term fairness goals. "Fair Game" provides a unique framework where the fairness goals can be adapted over time by only modifying the auditor and the different biases it quantifies. Thus,"Fair Game" aims to simulate the evolution of ethical and legal frameworks in the society by creating an auditor which sends feedback to a debiasing algorithm deployed around an ML system. This allows us to develop a flexible and adaptive-over-time framework to build Fair ML systems pre- and post-deployment.
References in corpus (15)
- Training language models to follow instructions with human feedback
- Equality of Opportunity in Supervised Learning
- Mastering Chess and Shogi by Self-Play with a General Reinforcement Learning Algorithm
- Fairness Testing: Testing Software for Discrimination
- Identifying and Correcting Label Bias in Machine Learning
- Aligning Large Language Models with Human: A Survey
- Null Compliance: NYC Local Law 144 and the Challenges of Algorithm Accountability
- Safe RLHF: Safe Reinforcement Learning from Human Feedback
- Fair Regression with Wasserstein Barycenters
- Survey on Fair Reinforcement Learning: Theory and Practice
- On the convergence of policy gradient methods to Nash equilibria in general stochastic games
- Algorithmic audits of algorithms, and the law
- Statistical Inference for Fairness Auditing
- On the Complexity of Differentially Private Best-Arm Identification with Fixed Confidence
- Risk-Sensitive Bayesian Games for Multi-Agent Reinforcement Learning under Policy Uncertainty