paper-with-me

홈 › Papers

AGR: Age Group fairness Reward for Bias Mitigation in LLMs

2024-09-06 · Shuirong Cao, Ruoxi Cheng, Zhiqiang Wang

LLMs can exhibit age biases, resulting in unequal treatment of individuals across age groups. While much research has addressed racial and gender biases, age bias remains little explored. The scarcity of instruction-tuning and preference datasets for age bias hampers its detection and measurement, and existing fine-tuning methods seldom address age-related fairness. In this paper, we construct age bias preference datasets and instruction-tuning datasets for RLHF. We introduce ARG, an age fairness reward to reduce differences in the response quality of LLMs across different age groups. Extensive experiments demonstrate that this reward significantly improves response accuracy and reduces performance disparities across age groups. Our source code and datasets are available at the anonymous \href{https://anonymous.4open.science/r/FairRLHF-D445/readme.md}{link}.

📄 PDF Abstract BibTeX arXiv:2409.04340

Code (0)

등록된 구현이 없습니다.

Tasks

Fairness

Similar Papers 제목 키워드 기반

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

2026-06-06 · Xiaoyan Zhao, Haoting Ni, Yang Zhang, Chunyuan Zheng 외 arxiv

Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models aim to capture such heterogeneity, they are often trained on imbalanc…

Towards Large Language Models that Benefit for All: Benchmarking Group Fairness in Reward Models

2025-03-10 · Kefan Song, Jin Yao, Runnan Jiang, Rohan Chandra 외

As Large Language Models (LLMs) become increasingly powerful and accessible to human users, ensuring fairness across diverse demographic groups, i.e., group fairness, is a critical ethical concern. However, current fairn…

AllBenchmarkingFairness

Software Fairness Dilemma: Is Bias Mitigation a Zero-Sum Game?

2025-08-05 · Zhenpeng Chen, Xinyue Li, Jie M. Zhang, Weisong Sun 외 arxiv

Fairness is a critical requirement for Machine Learning (ML) software, driving the development of numerous bias mitigation methods. Previous research has identified a leveling-down effect in bias mitigation for computer …

Bias Mitigation Post-processing for Individual and Group Fairness

2018-12-14 · Pranay K. Lohia, Karthikeyan Natesan Ramamurthy, Manish Bhide, Diptikalyan Saha 외

Whereas previous post-processing approaches for increasing the fairness of predictions of biased classifiers address only group fairness, we propose a method for increasing both individual and group fairness. Our novel f…

FairnessGeneral Classification

Fairness of ChatGPT

2023-05-22 · Yunqi Li, Lanjing Zhang, Yongfeng Zhang

Understanding and addressing unfairness in LLMs are crucial for responsible AI deployment. However, there is a limited number of quantitative analyses and in-depth studies regarding fairness evaluations in LLMs, especial…

Fairness