On the Interpretability of Quantum Neural Networks
arXiv:2308.11098 · doi:10.1007/s42484-024-00191-y
Abstract
Interpretability of artificial intelligence (AI) methods, particularly deep neural networks, is of great interest. This heightened focus stems from the widespread use of AI-backed systems. These systems, often relying on intricate neural architectures, can exhibit behavior that is challenging to explain and comprehend. The interpretability of such models is a crucial component of building trusted systems. Many methods exist to approach this problem, but they do not apply straightforwardly to the quantum setting. Here, we explore the interpretability of quantum neural networks using local model-agnostic interpretability measures commonly utilized for classical neural networks. Following this analysis, we generalize a classical technique called LIME, introducing Q-LIME, which produces explanations of quantum neural networks. A feature of our explanations is the delineation of the region in which data samples have been given a random label, likely subjects of inherently random quantum measurements. We view this as a step toward understanding how to build responsible and accountable quantum AI models.
References in corpus (6)
Cited by in corpus (7)
- Explaining Quantum Circuits with Shapley Values: Towards Explainable Quantum Machine Learning
- Networked Quantum Services
- Quantum Adjoint Convolutional Layers for Effective Data Representation
- QKAN: quantum Kolmogorov-Arnold networks with applications in machine learning and multivariate state preparation
- Engineering Quantum Reservoirs through Krylov Complexity, Expressivity and Observability
- IQNN-CS: Interpretable Quantum Neural Network for Credit Scoring
- Spectral analysis of the Koopman operator as a framework for recovering Hamiltonian parameters in open quantum systems