To tune or not to tune the number of trees in random forest?
arXiv:1705.05654
Abstract
The number of trees T in the random forest (RF) algorithm for supervised learning has to be set by the user. It is controversial whether T should simply be set to the largest computationally manageable value or whether a smaller T may in some cases be better. While the principle underlying bagging is that "more trees are better", in practice the classification error rate sometimes reaches a minimum before increasing again for increasing number of trees. The goal of this paper is four-fold: (i) providing theoretical results showing that the expected error rate may be a non-monotonous function of the number of trees and explaining under which circumstances this happens; (ii) providing theoretical results showing that such non-monotonous patterns cannot be observed for other performance measures such as the Brier score and the logarithmic loss (for classification) and the mean squared error (for regression); (iii) illustrating the extent of the problem through an application to a large number (n = 306) of datasets from the public database OpenML; (iv) finally arguing in favor of setting it to a computationally feasible large number, depending on convergence properties of the desired performance measure.
20 pages, 4 figures
Cited by in corpus (14)
- Automated Machine Learning: State-of-The-Art and Open Challenges
- The value of text for small business default prediction: A deep learning approach
- Oblique and rotation double random forest
- Two-Step Meta-Learning for Time-Series Forecasting Ensemble
- Randomization as Regularization: A Degrees of Freedom Explanation for Random Forest Success
- Collaborative Training of Balanced Random Forests for Open Set Domain Adaptation
- Experimental Investigation and Evaluation of Model-based Hyperparameter Optimization
- Continuous-Time Birth-Death MCMC for Bayesian Regression Tree Models
- Meta-Learning for Symbolic Hyperparameter Defaults
- Joints in Random Forests
- Best-scored Random Forest Classification
- Noise-Resilient Ensemble Learning using Evidence Accumulation Clustering
- Machine learning for automatic construction of pseudo-realistic pediatric abdominal phantoms
- A Dataset-Level Geometric Framework for Ensemble Classifiers