paper-with-me

홈 › Papers

Data Acquisition for Improving Model Fairness using Reinforcement Learning

2024-12-04 · Jahid Hasan, Romila Pradhan

Machine learning systems are increasingly being used in critical decision making such as healthcare, finance, and criminal justice. Concerns around their fairness have resulted in several bias mitigation techniques that emphasize the need for high-quality data to ensure fairer decisions. However, the role of earlier stages of machine learning pipelines in mitigating model bias has not been explored well. In this paper, we focus on the task of acquiring additional labeled data points for training the downstream machine learning model to rapidly improve its fairness. Since not all data points in a data pool are equally beneficial to the task of fairness, we generate an ordering in which data points should be acquired. We present DataSift, a data acquisition framework based on the idea of data valuation that relies on partitioning and multi-armed bandits to determine the most valuable data points to acquire. Over several iterations, DataSift selects a partition and randomly samples a batch of data points from the selected partition, evaluates the benefit of acquiring the batch on model fairness, and updates the utility of partitions depending on the benefit. To further improve the effectiveness and efficiency of evaluating batches, we leverage influence functions that estimate the effect of acquiring a batch without retraining the model. We empirically evaluate DataSift on several real-world and synthetic datasets and show that the fairness of a machine learning model can be significantly improved even while acquiring a few data points.

📄 PDF Abstract BibTeX arXiv:2412.03009

Code (0)

등록된 구현이 없습니다.

Tasks

Data ValuationFairnessMulti-Armed Banditsreinforcement-learningReinforcement Learning

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

FAL-CUR: Fair Active Learning using Uncertainty and Representativeness on Fair Clustering

2022-09-21 · Ricky Fajri, Akrati Saxena, Yulong Pei, Mykola Pechenizkiy

Active Learning (AL) techniques have proven to be highly effective in reducing data labeling costs across a range of machine learning tasks. Nevertheless, one known challenge of these methods is their potential to introd…

Active LearningClusteringFairness

Fair Knowledge Tracing in Second Language Acquisition

2024-12-23 · Weitao Tang, Guanliang Chen, Shuaishuai Zu, Jiangyi Luo

In second-language acquisition, predictive modeling aids educators in implementing diverse teaching strategies, attracting significant research attention. However, while model accuracy is widely explored, model fairness …

Deep LearningFairnessKnowledge TracingLanguage Acquisition

Intersectional Disentangling of Temporal and Acquisition Bias in Fetal Ultrasound

2026-05-01 · Aya Elgebaly, Joris Fournel, Benjamin Laine Jønch Jurgensen, Kamil Mikolaj 외 arxiv

Fairness studies of medical imaging AI often explain subgroup performance gaps through under-representation in the training data. We show that intersectional analysis can disentangle fairness and performance gaps arising…

Fairness in Reinforcement Learning

2016-11-09 · ICML 2017 8 · Shahin Jabbari, Matthew Joseph, Michael Kearns, Jamie Morgenstern 외

We initiate the study of fairness in reinforcement learning, where the actions of a learning algorithm may affect its environment and future rewards. Our fairness constraint requires that an algorithm never prefers one a…

Fairnessreinforcement-learningReinforcement LearningReinforcement Learning (RL)

Dynamic Fair Federated Learning Based on Reinforcement Learning

2023-11-02 · Weikang Chen, Junping Du, Yingxia Shao, Jia Wang 외

Federated learning enables a collaborative training and optimization of global models among a group of devices without sharing local data samples. However, the heterogeneity of data in federated learning can lead to unfa…

FairnessFederated Learningreinforcement-learningReinforcement Learning