paper-with-me

홈 › Papers

Toward Sufficient Statistical Power in Algorithmic Bias Assessment: A Test for ABROCA

2025-01-08 · Conrad Borchers

Algorithmic bias is a pressing concern in educational data mining (EDM), as it risks amplifying inequities in learning outcomes. The Area Between ROC Curves (ABROCA) metric is frequently used to measure discrepancies in model performance across demographic groups to quantify overall model fairness. However, its skewed distribution--especially when class or group imbalances exist--makes significance testing challenging. This study investigates ABROCA's distributional properties and contributes robust methods for its significance testing. Specifically, we address (1) whether ABROCA follows any known distribution, (2) how to reliably test for algorithmic bias using ABROCA, and (3) the statistical power achievable with ABROCA-based bias assessments under typical EDM sample specifications. Simulation results confirm that ABROCA does not match standard distributions, including those suited to accommodate skewness. We propose nonparametric randomization tests for ABROCA and demonstrate that reliably detecting bias with ABROCA requires large sample sizes or substantial effect sizes, particularly in imbalanced settings. Findings suggest that ABROCA-based bias evaluations based on sample sizes common in EDM tend to be underpowered, undermining the reliability of conclusions about model fairness. By offering open-source code to simulate power and statistically test ABROCA, this paper aims to foster more reliable statistical testing in EDM research. It supports broader efforts toward replicability and equity in educational modeling.

📄 PDF Abstract BibTeX arXiv:2501.04683

Code (0)

등록된 구현이 없습니다.

Tasks

Fairness

Similar Papers 제목 키워드 기반

A Simplicity Bubble Problem in Formal-Theoretic Learning Systems

2021-12-22 · Felipe S. Abrahão, Hector Zenil, Fabio Porto, Michael Winter 외

When mining large datasets in order to predict new data, limitations of the principles behind statistical machine learning pose a serious challenge not only to the Big Data deluge, but also to the traditional assumptions…

BIG-bench Machine Learning

Structural Gender Bias in Credit Scoring: Proxy Leakage

2026-01-26 · Navya SD, Sreekanth D, SS Uma Sankari arxiv

As financial institutions increasingly adopt machine learning for credit risk assessment, the persistence of algorithmic bias remains a critical barrier to equitable financial inclusion. This study provides a comprehensi…

Whither Bias Goes, I Will Go: An Integrative, Systematic Review of Algorithmic Bias Mitigation

2024-10-21 · Louis Hickman, Christopher Huynh, Jessica Gass, Brandon Booth 외

Machine learning (ML) models are increasingly used for personnel assessment and selection (e.g., resume screeners, automatically scored interviews). However, concerns have been raised throughout society that ML assessmen…

Fairness

Mitigating Bias in Algorithmic Hiring: Evaluating Claims and Practices

2019-06-21 · Manish Raghavan, Solon Barocas, Jon Kleinberg, Karen Levy

There has been rapidly growing interest in the use of algorithms in hiring, especially as a means to address or mitigate bias. Yet, to date, little is known about how these methods are used in practice. How are algorithm…

ABROCA Distributions For Algorithmic Bias Assessment: Considerations Around Interpretation

2024-11-28 · Conrad Borchers, Ryan S. Baker

Algorithmic bias continues to be a key concern of learning analytics. We study the statistical properties of the Absolute Between-ROC Area (ABROCA) metric. This fairness measure quantifies group-level differences in clas…

Fairness