Optimizing ensemble weights and hyperparameters of machine learning models for regression problems

Mohsen Shahhosseini; Guiping Hu; Hieu Pham

Machine Learning with Applications (Mar 2022)

Optimizing ensemble weights and hyperparameters of machine learning models for regression problems

Mohsen Shahhosseini,
Guiping Hu,
Hieu Pham

Affiliations

Mohsen Shahhosseini: Department of Industrial and Manufacturing Systems Engineering, Iowa State University, Ames, IA, 50011, USA
Guiping Hu: Department of Industrial and Manufacturing Systems Engineering, Iowa State University, Ames, IA, 50011, USA; Department of Sustainability, Golisano Institute for Sustainability, Rochester Institute of Technology, Rochester, NY 14623, USA; Corresponding author.
Hieu Pham: College of Business, The University of Alabama in Huntsville, Huntsville, AL 35899, USA

Journal volume & issue: Vol. 7
p. 100251

Abstract

Read online

Aggregating multiple learners through an ensemble of models aim to make better predictions by capturing the underlying distribution of the data more accurately. Different ensembling methods, such as bagging, boosting, and stacking/blending, have been studied and adopted extensively in research and practice. While bagging and boosting focus more on reducing variance and bias, respectively, stacking approaches target both by finding the optimal way to combine base learners. In stacking with the weighted average, ensembles are created from weighted averages of multiple base learners. It is known that tuning hyperparameters of each base learner inside the ensemble weight optimization process can produce better performing ensembles. To this end, an optimization-based nested algorithm that considers tuning hyperparameters as well as finding the optimal weights to combine ensembles (Generalized Weighted Ensemble with Internally Tuned Hyperparameters (GEM-ITH)) is designed. Besides, Bayesian search was used to speed-up the optimizing process and a heuristic was implemented to generate diverse and well-performing base learners. The algorithm is shown to be generalizable to real data sets through analyses with ten publicly available data sets.

Published in Machine Learning with Applications

ISSN: 2666-8270 (Online)
Publisher: Elsevier
Country of publisher: United Kingdom
LCC subjects: Science: Science (General): Cybernetics; Science: Mathematics: Instruments and machines: Electronic computers. Computer science
Website: https://www.journals.elsevier.com/machine-learning-with-applications

About the journal

Abstract

Keywords