paper-with-me

홈 › Papers

Optimally Weighted Ensembles of Regression Models: Exact Weight Optimization and Applications

2022-06-22 · Patrick Echtenbruck, Martina Echtenbruck, Joost Batenburg, Thomas Bäck, Boris Naujoks, Michael Emmerich

Automated model selection is often proposed to users to choose which machine learning model (or method) to apply to a given regression task. In this paper, we show that combining different regression models can yield better results than selecting a single ('best') regression model, and outline an efficient method that obtains optimally weighted convex linear combination from a heterogeneous set of regression models. More specifically, in this paper, a heuristic weight optimization, used in a preceding conference paper, is replaced by an exact optimization algorithm using convex quadratic programming. We prove convexity of the quadratic programming formulation for the straightforward formulation and for a formulation with weighted data points. The novel weight optimization is not only (more) exact but also more efficient. The methods we develop in this paper are implemented and made available via github-open source. They can be executed on commonly available hardware and offer a transparent and easy to interpret interface. The results indicate that the approach outperforms model selection methods on a range of data sets, including data sets with mixed variable type from drug discovery applications.

📄 PDF Abstract BibTeX arXiv:2206.11263

Code (0)

등록된 구현이 없습니다.

Tasks

Drug DiscoveryModel Selectionregression

Similar Papers 제목 키워드 기반

Kernel-based Optimally Weighted Conformal Prediction Intervals

2024-05-27 · JongHyeok Lee, Chen Xu, Yao Xie

In this work, we present a novel conformal prediction method for time-series, which we call Kernel-based Optimally Weighted Conformal Prediction Intervals (KOWCPI). Specifically, KOWCPI adapts the classic Reweighted Nada…

Conformal PredictionPredictionPrediction Intervalsquantile regression+2

On the Optimal Weighted $\ell_2$ Regularization in Overparameterized Linear Regression

2020-06-10 · NeurIPS 2020 12 · Denny Wu, Ji Xu

We consider the linear model $\mathbf{y} = \mathbf{X} \mathbf{\beta}_\star + \mathbf{\epsilon}$ with $\mathbf{X}\in \mathbb{R}^{n\times p}$ in the overparameterized regime $p>n$. We estimate $\mathbf{\beta}_\star$ via ge…

regression

No Free Lunch From Random Feature Ensembles

2024-12-06 · Benjamin S. Ruben, William L. Tong, Hamza Tahir Chaudhry, Cengiz Pehlevan

Given a budget on total model size, one must decide whether to train a single, large neural network or to combine the predictions of many smaller networks. We study this trade-off for ensembles of random-feature ridge re…

regression

Optimally Combining Classifiers Using Unlabeled Data

2015-03-05 · Akshay Balsubramani, Yoav Freund

We develop a worst-case analysis of aggregation of classifier ensembles for binary classification. The task of predicting to minimize error is formulated as a game played over a given set of unlabeled data (a transductiv…

Binary ClassificationGeneral Classification

One-shot Distributed Ridge Regression in High Dimensions

2020-01-01 · ICML 2020 1 · Yue Sheng, Edgar Dobriban

To scale up data analysis, distributed and parallel computing approaches are increasingly needed. Here we study a fundamental problem in this area: How to do ridge regression in a distributed computing environment? We st…

Distributed ComputingregressionUnityVocal Bursts Intensity Prediction