FAIR AI Models in High Energy Physics
arXiv:2212.05081 · doi:10.1088/2632-2153/ad12e3
Abstract
The findable, accessible, interoperable, and reusable (FAIR) data principles provide a framework for examining, evaluating, and improving how data is shared to facilitate scientific discovery. Generalizing these principles to research software and other digital products is an active area of research. Machine learning (ML) models -- algorithms that have been trained on data without being explicitly programmed -- and more generally, artificial intelligence (AI) models, are an important target for this because of the ever-increasing pace with which AI is transforming scientific domains, such as experimental high energy physics (HEP). In this paper, we propose a practical definition of FAIR principles for AI models in HEP and describe a template for the application of these principles. We demonstrate the template's use with an example AI model applied to HEP, in which a graph neural network is used to identify Higgs bosons decaying to two bottom quarks. We report on the robustness of this FAIR AI model, its portability across hardware architectures and software frameworks, and its interpretability.
34 pages, 9 figures, 10 tables
References in corpus (10)
- Observation of a new particle in the search for the Standard Model Higgs boson with the ATLAS detector at the LHC
- Observation of a new boson at a mass of 125 GeV with the CMS experiment at the LHC
- WaveNet: A Generative Model for Raw Audio
- OpenML: networked science in machine learning
- funcX: A Federated Function Serving Fabric for Science
- FAIR for AI: An interdisciplinary and international community building perspective
- A Detailed Study of Interpretability of Deep Neural Network based Top Taggers
- Snowmass 2021 Computational Frontier CompF03 Topical Group Report: Machine Learning
- Interpretable Geometric Deep Learning via Learnable Randomness Injection
- A Fresh Look at FAIR for Research Software