Towards Green Automated Machine Learning: Status Quo and Future Directions
arXiv:2111.05850 · doi:10.1613/jair.1.14340
Abstract
Automated machine learning (AutoML) strives for the automatic configuration of machine learning algorithms and their composition into an overall (software) solution - a machine learning pipeline - tailored to the learning task (dataset) at hand. Over the last decade, AutoML has developed into an independent research field with hundreds of contributions. At the same time, AutoML is being criticised for its high resource consumption as many approaches rely on the (costly) evaluation of many machine learning pipelines, as well as the expensive large scale experiments across many datasets and approaches. In the spirit of recent work on Green AI, this paper proposes Green AutoML, a paradigm to make the whole AutoML process more environmentally friendly. Therefore, we first elaborate on how to quantify the environmental footprint of an AutoML tool. Afterward, different strategies on how to design and benchmark an AutoML tool wrt. their "greenness", i.e. sustainability, are summarized. Finally, we elaborate on how to be transparent about the environmental footprint and what kind of research incentives could direct the community into a more sustainable AutoML research direction. Additionally, we propose a sustainability checklist to be attached to every AutoML paper featuring all core aspects of Green AutoML.
Published in Journal of Artificial Intelligence Research
References in corpus (27)
- Practical Bayesian Optimization of Machine Learning Algorithms
- Neural Architecture Search with Reinforcement Learning
- OpenML: networked science in machine learning
- NAS-Bench-101: Towards Reproducible Neural Architecture Search
- Non-stochastic Best Arm Identification and Hyperparameter Optimization
- Quantifying the Carbon Emissions of Machine Learning
- Carbon Emissions and Large Neural Network Training
- Freeze-Thaw Bayesian Optimization
- Carbontracker: Tracking and Predicting the Carbon Footprint of Training Deep Learning Models
- Multi-fidelity Bayesian Optimisation with Continuous Approximations
- Auto-WEKA: Combined Selection and Hyperparameter Optimization of Classification Algorithms
- NAS-Bench-201: Extending the Scope of Reproducible Neural Architecture Search
- Practical Multi-fidelity Bayesian Optimization for Hyperparameter Tuning
- Towards Green Automated Machine Learning: Status Quo and Future Directions
- HW-NAS-Bench:Hardware-Aware Neural Architecture Search Benchmark
- AutoML using Metadata Language Embeddings
- Multiple Adaptive Bayesian Linear Regression for Scalable Bayesian Optimization with Warm Start
- Zen-NAS: A Zero-Shot NAS for High-Performance Deep Image Recognition
- Multi-objective Asynchronous Successive Halving
- Automatic Termination for Hyperparameter Optimization
- Run2Survive: A Decision-theoretic Approach to Algorithm Selection based on Survival Analysis
- Fair and Green Hyperparameter Optimization via Multi-objective and Multiple Information Source Bayesian Optimization
- AutoML for Climate Change: A Call to Action
- IrEne: Interpretable Energy Prediction for Transformers
- Neural Model-based Optimization with Right-Censored Observations
- LeanML: A Design Pattern To Slash Avoidable Wastes in Machine Learning Projects
- Privileged Zero-Shot AutoML
Cited by in corpus (4)
- Towards Green Automated Machine Learning: Status Quo and Future Directions
- Estimating optical vegetation indices and biophysical variables for temperate forests with Sentinel-1 SAR data using machine learning techniques: A case study for Czechia
- Open and Sustainable AI: challenges, opportunities and the road ahead in the life sciences (October 2025 -- Version 2)
- Green AI: A systematic review and meta-analysis of its definitions, lifecycle models, hardware and measurement attempts