paper-with-me

홈 › Papers

Outperformance Score: A Universal Standardization Method for Confusion-Matrix-Based Classification Performance Metrics

2025-05-11 · Ningsheng Zhao, Trang Bui, Jia Yuan Yu, Krzysztof Dzieciolowski

Many classification performance metrics exist, each suited to a specific application. However, these metrics often differ in scale and can exhibit varying sensitivity to class imbalance rates in the test set. As a result, it is difficult to use the nominal values of these metrics to interpret and evaluate classification performances, especially when imbalance rates vary. To address this problem, we introduce the outperformance score function, a universal standardization method for confusion-matrix-based classification performance (CMBCP) metrics. It maps any given metric to a common scale of $[0,1]$, while providing a clear and consistent interpretation. Specifically, the outperformance score represents the percentile rank of the observed classification performance within a reference distribution of possible performances. This unified framework enables meaningful comparison and monitoring of classification performance across test sets with differing imbalance rates. We illustrate how the outperformance scores can be applied to a variety of commonly used classification performance metrics and demonstrate the robustness of our method through experiments on real-world datasets spanning multiple classification applications.

📄 PDF Abstract BibTeX arXiv:2505.07033

Code (0)

등록된 구현이 없습니다.

Tasks

Classification

Similar Papers 제목 키워드 기반

The MCC approaches the geometric mean of precision and recall as true negatives approach infinity

2023-04-30 · Jon Crall

The performance of a binary classifier is described by a confusion matrix with four entries: the number of true positives (TP), true negatives (TN), false positives (FP), and false negatives (FN). The Matthew's Correlati…

object-detectionObject Detection

Bridging the Gap: Unifying the Training and Evaluation of Neural Network Binary Classifiers

2020-09-02 · Nathan Tsoi, Kate Candon, Deyuan Li, Yofti Milkessa 외

While neural network binary classifiers are often evaluated on metrics such as Accuracy and $F_1$-Score, they are commonly trained with a cross-entropy objective. How can this training-evaluation gap be addressed? While …

General Classification

AnyLoss: Transforming Classification Metrics into Loss Functions

2024-05-23 · Doheon Han, Nuno Moniz, Nitesh V Chawla

Many evaluation metrics can be used to assess the performance of models in binary classification tasks. However, most of them are derived from a confusion matrix in a non-differentiable form, making it very difficult to …

Binary ClassificationClassificationModel Selection

Aligning Multiclass Neural Network Classifier Criterion with Task Performance via $F_β$-Score

2024-05-31 · Nathan Tsoi, Deyuan Li, Taesoo Daniel Lee, Marynel Vázquez

Multiclass neural network classifiers are typically trained using cross-entropy loss. Following training, the performance of this same neural network is evaluated using an application-specific metric based on the multicl…

Binary Classification

Prayatul Matrix: A Direct Comparison Approach to Evaluate Performance of Supervised Machine Learning Models

2022-09-26 · Anupam Biswas

Performance comparison of supervised machine learning (ML) models are widely done in terms of different confusion matrix based scores obtained on test datasets. However, a dataset comprises several instances having diffe…