Hyperparameter optimization: Foundations, algorithms, best practices, and open challenges

www.lmu.de | UB | Blättern | Hilfe

Zur erweiterten Suche

English

Zur erweiterten Suche

Bischl, Bernd ORCID: https://orcid.org/0000-0001-6002-6980; Binder, Martin; Lang, Michel ORCID: https://orcid.org/0000-0001-9754-0393; Pielok, Tobias; Richter, Jakob ORCID: https://orcid.org/0000-0003-4481-5554; Coors, Stefan ORCID: https://orcid.org/0000-0002-7465-2146; Thomas, Janek; Ullmann, Theresa ORCID: https://orcid.org/0000-0003-1215-8561; Becker, Marc ORCID: https://orcid.org/0000-0002-8115-0400; Boulesteix, Anne‐Laure ORCID: https://orcid.org/0000-0002-2729-0947; Deng, Difan und Lindauer, Marius ORCID: https://orcid.org/0000-0002-9675-3175 (2023): Hyperparameter optimization: Foundations, algorithms, best practices, and open challenges. In: WIREs Data Mining and Knowledge Discovery, Bd. 13, Nr. 2 [PDF, 6MB]

[thumbnail of WIREs_Data_Min___Knowl_-_2023_-_Bischl_-_Hyperparameter_optimization__Foundations__algorithms__best_practices__and_open.pdf]

Vorschau

Creative Commons: Namensnennung 4.0 (CC-BY)

DOI: 10.1002/widm.1484

Abstract

Most machine learning algorithms are configured by a set of hyperparameters whose values must be carefully chosen and which often considerably impact performance. To avoid a time-consuming and irreproducible manual process of trial-and-error to find well-performing hyperparameter configurations, various automatic hyperparameter optimization (HPO) methods—for example, based on resampling error estimation for supervised machine learning—can be employed. After introducing HPO from a general perspective, this paper reviews important HPO methods, from simple techniques such as grid or random search to more advanced methods like evolution strategies, Bayesian optimization, Hyperband, and racing. This work gives practical recommendations regarding important choices to be made when conducting HPO, including the HPO algorithms themselves, performance evaluation, how to combine HPO with machine learning pipelines, runtime improvements, and parallelization.

This article is categorized under:

Algorithmic Development > Statistics Technologies > Machine Learning Technologies > Prediction

Dokumententyp:	Zeitschriftenartikel
Fakultät:	Mathematik, Informatik und Statistik > Statistik
Themengebiete:	500 Naturwissenschaften und Mathematik > 510 Mathematik
URN:	urn:nbn:de:bvb:19-epub-108819-7
ISSN:	1942-4787
Sprache:	Englisch
Dokumenten ID:	108819
Datum der Veröffentlichung auf Open Access LMU:	13. Mrz. 2024 10:39
Letzte Änderungen:	13. Mrz. 2024 10:39

Dokument bearbeiten