paper-with-me

Papers

Ensemble feature selection with clustering for analysis of high-dimensional, correlated clinical data in the search for Alzheimer's disease biomarkers

2022-07-06 · Annette Spooner, Gelareh Mohammadi, Perminder S. Sachdev, Henry Brodaty, Arcot Sowmya

Healthcare datasets often contain groups of highly correlated features, such as features from the same biological system. When feature selection is applied to these datasets to identify the most important features, the biases inherent in some multivariate feature selectors due to correlated features make it difficult for these methods to distinguish between the important and irrelevant features and the results of the feature selection process can be unstable. Feature selection ensembles, which aggregate the results of multiple individual base feature selectors, have been investigated as a means of stabilising feature selection results, but do not address the problem of correlated features. We present a novel framework to create feature selection ensembles from multivariate feature selectors while taking into account the biases produced by groups of correlated features, using agglomerative hierarchical clustering in a pre-processing step. These methods were applied to two real-world datasets from studies of Alzheimer's disease (AD), a progressive neurodegenerative disease that has no cure and is not yet fully understood. Our results show a marked improvement in the stability of features selected over the models without clustering, and the features selected by these models are in keeping with the findings in the AD literature.

📄 PDF Abstract BibTeX arXiv:2207.02380

Code (0)

등록된 구현이 없습니다.

Tasks

Clusteringfeature selection

Methods 이 논문이 사용한 방법론

Feature Selection Feature selection, also known as variable selection, attribute selection or variable subset selection, is the process of selecting a subset of relevant features (variables,…
BASE 설명 없음

Similar Papers 제목 키워드 기반

ESDF: Ensemble Selection using Diversity and Frequency

2015-08-18 · Shouvick Mondal, Arko Banerjee

Recently ensemble selection for consensus clustering has emerged as a research problem in Machine Intelligence. Normally consensus clustering algorithms take into account the entire ensemble of clustering, where there is…

ClusteringDiversity

Deep Learning and Linear Programming for Automated Ensemble Forecasting and Interpretation

2022-01-02 · Lars Lien Ankile, Kjartan Krange

This paper presents an ensemble forecasting method that shows strong results on the M4 Competition dataset by decreasing feature and model selection assumptions, termed DONUT (DO Not UTilize human beliefs). Our assumptio…

Feature ImportanceModel SelectionTime SeriesTime Series Analysis

On Hyperparameter Search in Cluster Ensembles

2018-03-29 · Luzie Helfmann, Johannes von Lindheim, Mattes Mollenhauer, Ralf Banisch

Quality assessments of models in unsupervised learning and clustering verification in particular have been a long-standing problem in the machine learning research. The lack of robust and universally applicable cluster v…

ClusteringClustering Ensemble

Statistical Parameter Selection for Clustering Persistence Diagrams

2019-10-17 · Max Kontak, Jules Vidal, Julien Tierny

In urgent decision making applications, ensemble simulations are an important way to determine different outcome scenarios based on currently available data. In this paper, we will analyze the output of ensemble simulati…

ClusteringDecision Making

Greedy Feature Selection for Subspace Clustering

2013-03-19 · Eva L. Dyer, Aswin C. Sankaranarayanan, Richard G. Baraniuk

Unions of subspaces provide a powerful generalization to linear subspace models for collections of high-dimensional data. To learn a union of subspaces from a collection of data, sets of signals in the collection that be…

Clusteringfeature selection