paper-with-me

홈 › Papers

FIB: A Method for Evaluation of Feature Impact Balance in Multi-Dimensional Data

2022-07-10 · Xavier F. Cadet, Sara Ahmadi-Abhari, Hamed Haddadi

Errors might not have the same consequences depending on the task at hand. Nevertheless, there is limited research investigating the impact of imbalance in the contribution of different features in an error vector. Therefore, we propose the Feature Impact Balance (FIB) score. It measures whether there is a balanced impact of features in the discrepancies between two vectors. We designed the FIB score to lie in [0, 1]. Scores close to 0 indicate that a small number of features contribute to most of the error, and scores close to 1 indicate that most features contribute to the error equally. We experimentally study the FIB on different datasets, using AutoEncoders and Variational AutoEncoders. We show how the feature impact balance varies during training and showcase its usability to support model selection for single output and multi-output tasks.

📄 PDF Abstract BibTeX arXiv:2207.04500

Code (0)

등록된 구현이 없습니다.

Tasks

Model Selection

Similar Papers 제목 키워드 기반

MH-FSF: A Unified Framework for Overcoming Benchmarking and Reproducibility Limitations in Feature Selection Evaluation

2025-07-11 · Vanderson Rocha, Diego Kreutz, Gabriel Canto, Hendrio Bragança 외 arxiv

Feature selection is vital for building effective predictive models, as it reduces dimensionality and emphasizes key features. However, current research often suffers from limited benchmarking and reliance on proprietary…

Malware Detection

Feature Dimensionality Outweighs Model Complexity in Breast Cancer Subtype Classification Using TCGA-BRCA Gene Expression Data

2026-05-07 · Meena Al Hasani arxiv

Accurate classification of breast cancer subtypes from gene expression data is critical for diagnosis and treatment selection. However, such datasets are characterized by high dimensionality and limited sample size, posi…

An Empirical Evaluation of the t-SNE Algorithm for Data Visualization in Structural Engineering

2021-09-18 · Parisa Hajibabaee, Farhad Pourkamali-Anaraki, Mohammad Amin Hariri-Ardebili

A fundamental task in machine learning involves visualizing high-dimensional data sets that arise in high-impact application domains. When considering the context of large imbalanced data, this problem becomes much more …

Data Visualization

Graph Neural Network-Driven Hierarchical Mining for Complex Imbalanced Data

2025-02-06 · Yijiashun Qi, Quanchao Lu, Shiyu Dou, Xiaoxuan Sun 외

This study presents a hierarchical mining framework for high-dimensional imbalanced data, leveraging a depth graph model to address the inherent performance limitations of conventional approaches in handling complex, hig…

Graph Neural Network

New Hard-thresholding Rules based on Data Splitting in High-dimensional Imbalanced Classification

2021-11-05 · Arezou Mojiri, Abbas Khalili, Ali Zeinal Hamadani

In binary classification, imbalance refers to situations in which one class is heavily under-represented. This issue is due to either a data collection process or because one class is indeed rare in a population. Imbalan…

Binary Classificationimbalanced classification