Towards a learning-based performance modeling for accelerating Deep Neural Networks
arXiv:2212.05031 · doi:10.1007/978-3-030-24289-3_49
Abstract
Emerging applications such as Deep Learning are often data-driven, thus traditional approaches based on auto-tuners are not performance effective across the wide range of inputs used in practice. In the present paper, we start an investigation of predictive models based on machine learning techniques in order to optimize Convolution Neural Networks (CNNs). As a use-case, we focus on the ARM Compute Library which provides three different implementations of the convolution operator at different numeric precision. Starting from a collation of benchmarks, we build and validate models learned by Decision Tree and naive Bayesian classifier. Preliminary experiments on Midgard-based ARM Mali GPU show that our predictive model outperforms all the convolution operators manually selected by the library.
References in corpus (1)
Cited by in corpus (6)
- Binary classification of proteins by a Machine Learning approach
- Skin Cancer Classification using Inception Network and Transfer Learning
- Implementing a scalable and elastic computing environment based on Cloud Containers
- A new method for binary classification of proteins with Machine Learning
- High Performance Computing and Computational Intelligence Applications with MultiChaos Perspective
- IoT to monitor people flow in areas of public interest