paper-with-me

Papers

Efficient Estimation of Generalization Error and Bias-Variance Components of Ensembles

2017-11-15 · Dhruv Mahajan, Vivek Gupta, S. Sathiya Keerthi, Sellamanickam Sundararajan, Shravan Narayanamurthy, Rahul Kidambi

For many applications, an ensemble of base classifiers is an effective solution. The tuning of its parameters(number of classes, amount of data on which each classifier is to be trained on, etc.) requires G, the generalization error of a given ensemble. The efficient estimation of G is the focus of this paper. The key idea is to approximate the variance of the class scores/probabilities of the base classifiers over the randomness imposed by the training subset by normal/beta distribution at each point x in the input feature space. We estimate the parameters of the distribution using a small set of randomly chosen base classifiers and use those parameters to give efficient estimation schemes for G. We give empirical evidence for the quality of the various estimators. We also demonstrate their usefulness in making design choices such as the number of classifiers in the ensemble and the size of a subset of data used for training that is needed to achieve a certain value of generalization error. Our approach also has great potential for designing distributed ensemble classifiers.

📄 PDF Abstract BibTeX arXiv:1711.05482

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Understanding the Generalization of Bilevel Programming in Hyperparameter Optimization: A Tale of Bias-Variance Decomposition

2026-02-20 · Yubo Zhou, Jun Shu, Junmin Liu, Deyu Meng arxiv

Gradient-based hyperparameter optimization (HPO) have emerged recently, leveraging bilevel programming techniques to optimize hyperparameter by estimating hypergradient w.r.t. validation loss. Nevertheless, previous theo…

Hyperparameter OptimizationFew-Shot Learning

A Generalized Bias-Variance Decomposition for Bregman Divergences

2025-11-11 · David Pfau arxiv

The bias-variance decomposition is a central result in statistics and machine learning, but is typically presented only for the squared error. We present a generalization of the bias-variance decomposition where the pred…

Understanding Generalization in Adversarial Training via the Bias-Variance Decomposition

2021-03-17 · Yaodong Yu, Zitong Yang, Edgar Dobriban, Jacob Steinhardt 외

Adversarially trained models exhibit a large generalization gap: they can interpolate the training set even for large perturbation radii, but at the cost of large test error on clean samples. To investigate this gap, we …

On the Inter-relationships among Drift rate, Forgetting rate, Bias/variance profile and Error

2018-01-29 · Nayyar A. Zaidi, Geoffrey I. Webb, Francois Petitjean, Germain Forestier

We propose two general and falsifiable hypotheses about expectations on generalization error when learning in the context of concept drift. One posits that as drift rate increases, the forgetting rate that minimizes gene…

Rethink the Connections among Generalization, Memorization and the Spectral Bias of DNNs

2020-04-29 · Xiao Zhang, Haoyi Xiong, Dongrui Wu

Over-parameterized deep neural networks (DNNs) with sufficient capacity to memorize random noise can achieve excellent generalization performance, challenging the bias-variance trade-off in classical learning theory. Rec…

Learning TheoryMemorization