On the Definition of Appropriate Trust and the Tools that Come with it
arXiv:2309.11937 · doi:10.1109/CSCE60160.2023.00256
Abstract
Evaluating the efficiency of human-AI interactions is challenging, including subjective and objective quality aspects. With the focus on the human experience of the explanations, evaluations of explanation methods have become mostly subjective, making comparative evaluations almost impossible and highly linked to the individual user. However, it is commonly agreed that one aspect of explanation quality is how effectively the user can detect if the predictions are trustworthy and correct, i.e., if the explanations can increase the user's appropriate trust in the model. This paper starts with the definitions of appropriate trust from the literature. It compares the definitions with model performance evaluation, showing the strong similarities between appropriate trust and model performance evaluation. The paper's main contribution is a novel approach to evaluating appropriate trust by taking advantage of the likenesses between definitions. The paper offers several straightforward evaluation methods for different aspects of user performance, including suggesting a method for measuring uncertainty and appropriate trust in regression.
8 pages, 3 figures, Conference: ICDATA 2023
References in corpus (5)
- Towards A Rigorous Science of Interpretable Machine Learning
- Opportunities and Challenges in Explainable Artificial Intelligence (XAI): A Survey
- Proxy Tasks and Subjective Measures Can Be Misleading in Evaluating Explainable AI Systems
- Post-hoc explanation of black-box classifiers using confident itemsets
- A Meta Survey of Quality Evaluation Criteria in Explanation Methods