The Impact of Explanations on Fairness in Human-AI Decision-Making: Protected vs Proxy Features
arXiv:2310.08617 · doi:10.1145/3640543.3645210
Abstract
AI systems have been known to amplify biases in real-world data. Explanations may help human-AI teams address these biases for fairer decision-making. Typically, explanations focus on salient input features. If a model is biased against some protected group, explanations may include features that demonstrate this bias, but when biases are realized through proxy features, the relationship between this proxy feature and the protected one may be less clear to a human. In this work, we study the effect of the presence of protected and proxy features on participants' perception of model fairness and their ability to improve demographic parity over an AI alone. Further, we examine how different treatments -- explanations, model bias disclosure and proxy correlation disclosure -- affect fairness perception and parity. We find that explanations help people detect direct but not indirect biases. Additionally, regardless of bias type, explanations tend to increase agreement with model biases. Disclosures can help mitigate this effect for indirect biases, improving both unfairness recognition and decision-making fairness. We hope that our findings can help guide further research into advancing explanations in support of fair human-AI decision-making.
IUI 2024
References in corpus (10)
- Towards A Rigorous Science of Interpretable Machine Learning
- To Trust or to Think: Cognitive Forcing Functions Can Reduce Overreliance on AI in AI-assisted Decision-making
- Effect of Confidence and Explanation on Accuracy and Trust Calibration in AI-Assisted Decision Making
- 'It's Reducing a Human Being to a Percentage'; Perceptions of Justice in Algorithmic Decisions
- What Do We Want From Explainable Artificial Intelligence (XAI)? -- A Stakeholder Perspective on XAI and a Conceptual Model Guiding Interdisciplinary XAI Research
- Proxy Tasks and Subjective Measures Can Be Misleading in Evaluating Explainable AI Systems
- Explaining Models: An Empirical Study of How Explanations Impact Fairness Judgment
- Do People Engage Cognitively with AI? Impact of AI Assistance on Incidental Learning
- Understanding the Effect of Out-of-distribution Examples and Interactive Explanations on Human-AI Decision Making
- Misplaced Trust: Measuring the Interference of Machine Learning in Human Decision-Making