paper-with-me

홈 › Papers

The Heterogeneous Ensembles of Standard Classification Algorithms (HESCA): the Whole is Greater than the Sum of its Parts

2017-10-25 · James Large, Jason Lines, Anthony Bagnall

Building classification models is an intrinsically practical exercise that requires many design decisions prior to deployment. We aim to provide some guidance in this decision making process. Specifically, given a classification problem with real valued attributes, we consider which classifier or family of classifiers should one use. Strong contenders are tree based homogeneous ensembles, support vector machines or deep neural networks. All three families of model could claim to be state-of-the-art, and yet it is not clear when one is preferable to the others. Our extensive experiments with over 200 data sets from two distinct archives demonstrate that, rather than choose a single family and expend computing resources on optimising that model, it is significantly better to build simpler versions of classifiers from each family and ensemble. We show that the Heterogeneous Ensembles of Standard Classification Algorithms (HESCA), which ensembles based on error estimates formed on the train data, is significantly better (in terms of error, balanced error, negative log likelihood and area under the ROC curve) than its individual components, picking the component that is best on train data, and a support vector machine tuned over 1089 different parameter configurations. We demonstrate HESCA+, which contains a deep neural network, a support vector machine and two decision tree forests, is significantly better than its components, picking the best component, and HESCA. We analyse the results further and find that HESCA and HESCA+ are of particular value when the train set size is relatively small and the problem has multiple classes. HESCA is a fast approach that is, on average, as good as state-of-the-art classifiers, whereas HESCA+ is significantly better than average and represents a strong benchmark for future research.

📄 PDF Abstract BibTeX arXiv:1710.09220

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingGeneral Classification

Similar Papers 제목 키워드 기반

A Large-Scale Benchmark of Cross-Modal Learning for Histology and Gene Expression in Spatial Transcriptomics

2025-08-02 · Rushin H. Gindra, Giovanni Palla, Mathias Nguyen, Sophia J. Wagner 외 arxiv

Spatial transcriptomics enables simultaneous measurement of gene expression and tissue morphology, offering unprecedented insights into cellular organization and disease mechanisms. However, the field lacks comprehensive…

Pooling homogeneous ensembles to build heterogeneous ones

2018-02-21 · Maryam Sabzevari, Gonzalo Martínez-Muñoz, Alberto Suárez

In ensemble methods, the outputs of a collection of diverse classifiers are combined in the expectation that the global prediction be more accurate than the individual ones. Heterogeneous ensembles consist of predictors …

Pathologies of Predictive Diversity in Deep Ensembles

2023-02-01 · Taiga Abe, E. Kelly Buchanan, Geoff Pleiss, John P. Cunningham

Classic results establish that encouraging predictive diversity improves performance in ensembles of low-capacity models, e.g. through bagging or boosting. Here we demonstrate that these intuitions do not apply to high-c…

Diversity

Developing parsimonious ensembles using predictor diversity within a reinforcement learning framework

2021-02-15 · Ana Stanescu, Gaurav Pandey

Heterogeneous ensembles that can aggregate an unrestricted number and variety of base predictors can effectively address challenging prediction problems. In particular, accurate ensembles that are also parsimonious, i.e.…

Diversityreinforcement-learningReinforcement LearningReinforcement Learning (RL)+1

Systematic Ensemble Learning for Regression

2014-03-28 · Roberto Aldave, Jean-Pierre Dussault

The motivation of this work is to improve the performance of standard stacking approaches or ensembles, which are composed of simple, heterogeneous base models, through the integration of the generation and selection sta…

Ensemble Learningregression