paper-with-me

Papers

Default Machine Learning Hyperparameters Do Not Provide Informative Initialization for Bayesian Optimization

2026-02-09 · Nicolás Villagrán Prieto, Eduardo C. Garrido-Merchán arxiv

Bayesian Optimization (BO) is a standard tool for hyperparameter tuning thanks to its sample efficiency on expensive black-box functions. While most BO pipelines begin with uniform random initialization, default hyperparameter values shipped with popular ML libraries such as scikit-learn encode implicit expert knowledge and could serve as informative starting points that accelerate convergence. This hypothesis, despite its intuitive appeal, has remained largely unexamined. We formalize the idea by initializing BO with points drawn from truncated Gaussian distributions centered at library defaults and compare the resulting trajectories against a uniform-random baseline. We conduct an extensive empirical evaluation spanning three BO back-ends (BoTorch, Optuna, Scikit-Optimize), three model families (Random Forests, Support Vector Machines, Multilayer Perceptrons), and five benchmark datasets covering classification and regression tasks. Performance is assessed through convergence speed and final predictive quality, and statistical significance is determined via one-sided binomial tests. Across all conditions, default-informed initialization yields no statistically significant advantage over purely random sampling, with p-values ranging from 0.141 to 0.908. A sensitivity analysis on the prior variance confirms that, while tighter concentration around the defaults improves early evaluations, this transient benefit vanishes as optimization progresses, leaving final performance unchanged. Our results provide no evidence that default hyperparameters encode useful directional information for optimization. We therefore recommend that practitioners treat hyperparameter tuning as an integral part of model development and favor principled, data-driven search strategies over heuristic reliance on library defaults.

📄 PDF Abstract BibTeX arXiv:2602.08774

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Tunability: Importance of Hyperparameters of Machine Learning Algorithms

2018-02-26 · Philipp Probst, Bernd Bischl, Anne-Laure Boulesteix

Modern supervised machine learning algorithms involve hyperparameters that have to be set before running them. Options for setting hyperparameters are default values from the software package, manual configuration by the…

BenchmarkingBIG-bench Machine Learning

Importance of Tuning Hyperparameters of Machine Learning Algorithms

2020-07-15 · Hilde J. P. Weerts, Andreas C. Mueller, Joaquin Vanschoren

The performance of many machine learning algorithms depends on their hyperparameter settings. The goal of this study is to determine whether it is important to tune a hyperparameter or whether it can be safely set to a d…

BIG-bench Machine Learning

Informative Initialization and Kernel Selection Improves t-SNE for Biological Sequences

2022-11-16 · Prakash Chourasia, Sarwan Ali, Murray Patterson

The t-distributed stochastic neighbor embedding (t- SNE) is a method for interpreting high dimensional (HD) data by mapping each point to a low dimensional (LD) space (usually two-dimensional). It seeks to retain the str…

It's Much Easier for Neural Networks to learn Game of Life Dynamics with the Right Activation Function: Polynomial Kolmogorov-Arnold Networks

2026-06-22 · Tashin Ahmed, Q. Tyrell Davis arxiv

Previous work has found a gap between the scale of neural networks that reliably learn Conway's Game of Life, and minimal networks capable of representing the classic cellular automaton with hard-coded parameter values. …

Interpretable Machine Learning

Rethinking Default Values: a Low Cost and Efficient Strategy to Define Hyperparameters

2020-07-31 · Rafael Gomes Mantovani, André Luis Debiaso Rossi, Edesio Alcobaça, Jadson Castro Gertrudes 외

Machine Learning (ML) algorithms have been increasingly applied to problems from several different areas. Despite their growing popularity, their predictive performance is usually affected by the values assigned to their…