paper-with-me

Papers

An Empirical Characterization of Fair Machine Learning For Clinical Risk Prediction

2020-07-20 · Stephen R. Pfohl, Agata Foryciarz, Nigam H. Shah

The use of machine learning to guide clinical decision making has the potential to worsen existing health disparities. Several recent works frame the problem as that of algorithmic fairness, a framework that has attracted considerable attention and criticism. However, the appropriateness of this framework is unclear due to both ethical as well as technical considerations, the latter of which include trade-offs between measures of fairness and model performance that are not well-understood for predictive models of clinical outcomes. To inform the ongoing debate, we conduct an empirical study to characterize the impact of penalizing group fairness violations on an array of measures of model performance and group fairness. We repeat the analyses across multiple observational healthcare databases, clinical outcomes, and sensitive attributes. We find that procedures that penalize differences between the distributions of predictions across groups induce nearly-universal degradation of multiple performance metrics within groups. On examining the secondary impact of these procedures, we observe heterogeneity of the effect of these procedures on measures of fairness in calibration and ranking across experimental conditions. Beyond the reported trade-offs, we emphasize that analyses of algorithmic fairness in healthcare lack the contextual grounding and causal awareness necessary to reason about the mechanisms that lead to health disparities, as well as about the potential of algorithmic fairness methods to counteract those mechanisms. In light of these limitations, we encourage researchers building predictive models for clinical use to step outside the algorithmic fairness frame and engage critically with the broader sociotechnical context surrounding the use of machine learning in healthcare.

📄 PDF Abstract BibTeX arXiv:2007.10306

Code (1)

som-shahlab/fairness_benchmark 공식 구현 pytorch

Tasks

BIG-bench Machine LearningDecision MakingFairness

Similar Papers 제목 키워드 기반

When Personalization Harms: Reconsidering the Use of Group Attributes in Prediction

2022-06-04 · Vinith M. Suriyakumar, Marzyeh Ghassemi, Berk Ustun

Machine learning models are often personalized with categorical attributes that are protected, sensitive, self-reported, or costly to acquire. In this work, we show models that are personalized with group attributes can …

Evaluating Intersectional Fairness across Clinical Machine Learning Use Cases using Fairlogue and the All of Us Research Program

2026-04-07 · Nick Souligne, Vignesh Subbian arxiv

Intersectional biases in healthcare data can produce compound disparities in clinical machine learning models, yet most fairness evaluations assess demographic attributes independently. FairLogue, a toolkit for intersect…

Equitable Survival Prediction: A Fairness-Aware Survival Modeling (FASM) Approach

2025-10-23 · Mingxuan Liu, Yilin Ning, Haoyuan Wang, Chuan Hong 외 arxiv

As machine learning models become increasingly integrated into healthcare, structural inequities and social biases embedded in clinical data can be perpetuated or even amplified by data-driven models. In survival analysi…

An empirical evaluation of the risks of AI model updates using clinical data: stability, arbitrariness, and fairness

2026-04-27 · Ioannis Bilionis, Ricardo C. Berrios, Luis Fernandez-Luque, Carlos Castillo arxiv

Artificial Intelligence (AI) and Machine Learning (ML) models used in clinical settings are increasingly deployed to support clinical decision-making. However, when training data become stale due to changes in demographi…

The Effect of Enforcing Fairness on Reshaping Explanations in Machine Learning Models

2025-12-01 · Joshua Wolff Anderson, Shyam Visweswaran arxiv

Trustworthy machine learning in healthcare requires strong predictive performance, fairness, and explanations. While it is known that improving fairness can affect predictive performance, little is known about how fairne…

Feature Importance