paper-with-me

Papers

Examining Imbalance Effects on Performance and Demographic Fairness of Clinical Language Models

2024-12-23 · Precious Jones, Weisi Liu, I-Chan Huang, Xiaolei Huang

Data imbalance is a fundamental challenge in applying language models to biomedical applications, particularly in ICD code prediction tasks where label and demographic distributions are uneven. While state-of-the-art language models have been increasingly adopted in biomedical tasks, few studies have systematically examined how data imbalance affects model performance and fairness across demographic groups. This study fills the gap by statistically probing the relationship between data imbalance and model performance in ICD code prediction. We analyze imbalances in a standard benchmark data across gender, age, ethnicity, and social determinants of health by state-of-the-art biomedical language models. By deploying diverse performance metrics and statistical analyses, we explore the influence of data imbalance on performance variations and demographic fairness. Our study shows that data imbalance significantly impacts model performance and fairness, but feature similarity to the majority class may be a more critical factor. We believe this study provides valuable insights for developing more equitable and robust language models in healthcare applications.

📄 PDF Abstract BibTeX arXiv:2412.17803

Code (0)

등록된 구현이 없습니다.

Tasks

Fairness

Similar Papers 제목 키워드 기반

Fairness Index Measures to Evaluate Bias in Biometric Recognition

2023-06-19 · Ketan Kotwal, Sebastien Marcel

The demographic disparity of biometric systems has led to serious concerns regarding their societal impact as well as applicability of such systems in private and public domains. A quantitative evaluation of demographic …

BenchmarkingFairness

Toward Fair Federated Learning under Demographic Disparities and Data Imbalance

2025-05-14 · Qiming Wu, Siqi Li, Doudou Zhou, Nan Liu

Ensuring fairness is critical when applying artificial intelligence to high-stakes domains such as healthcare, where predictive models trained on imbalanced and demographically skewed data risk exacerbating existing disp…

FairnessFederated LearningPrivacy Preserving

A Comparison of Differential Performance Metrics for the Evaluation of Automatic Speaker Verification Fairness

2024-04-27 · Oubaida Chouchane, Christoph Busch, Chiara Galdi, Nicholas Evans 외

When decisions are made and when personal data is treated by automated processes, there is an expectation of fairness -- that members of different demographic groups receive equitable treatment. This expectation applies …

Face RecognitionFairnessSpeaker Verification

Fairness on Synthetic Visual and Thermal Mask Images

2022-09-19 · Kenneth Lai, Vlad Shmerko, Svetlana Yanushkevich

In this paper, we study performance and fairness on visual and thermal images and expand the assessment to masked synthetic images. Using the SpeakingFace and Thermal-Mask dataset, we propose a process to assess fairness…

DiversityFairness

LLMs on Trial: Evaluating Judicial Fairness for Large Language Models

2025-07-14 · Yiran Hu, Zongyue Xue, Haitao Li, Siyuan Zheng 외 arxiv

Large Language Models (LLMs) are increasingly used in high-stakes fields where their decisions impact rights and equity. However, LLMs' judicial fairness and implications for social justice remain underexplored. When LLM…