paper-with-me

홈 › Papers

Bias Neutralization Framework: Measuring Fairness in Large Language Models with Bias Intelligence Quotient (BiQ)

2024-04-28 · Malur Narayan, John Pasmore, Elton Sampaio, Vijay Raghavan, Gabriella Waters

The burgeoning influence of Large Language Models (LLMs) in shaping public discourse and decision-making underscores the imperative to address inherent biases within these AI systems. In the wake of AI's expansive integration across sectors, addressing racial bias in LLMs has never been more critical. This paper introduces a novel framework called Comprehensive Bias Neutralization Framework (CBNF) which embodies an innovative approach to quantifying and mitigating biases within LLMs. Our framework combines the Large Language Model Bias Index (LLMBI) [Oketunji, A., Anas, M., Saina, D., (2023)] and Bias removaL with No Demographics (BLIND) [Orgad, H., Belinkov, Y. (2023)] methodologies to create a new metric called Bias Intelligence Quotient (BiQ)which detects, measures, and mitigates racial bias in LLMs without reliance on demographic annotations. By introducing a new metric called BiQ that enhances LLMBI with additional fairness metrics, CBNF offers a multi-dimensional metric for bias assessment, underscoring the necessity of a nuanced approach to fairness in AI [Mehrabi et al., 2021]. This paper presents a detailed analysis of Latimer AI (a language model incrementally trained on black history and culture) in comparison to ChatGPT 3.5, illustrating Latimer AI's efficacy in detecting racial, cultural, and gender biases through targeted training and refined bias mitigation strategies [Latimer & Bender, 2023].

📄 PDF Abstract BibTeX arXiv:2404.18276

Code (0)

등록된 구현이 없습니다.

Tasks

Decision MakingFairnessLanguage ModelingLanguage ModellingLarge Language Model

Similar Papers 제목 키워드 기반

Fairness via Representation Neutralization

2021-06-23 · NeurIPS 2021 12 · Mengnan Du, Subhabrata Mukherjee, Guanchu Wang, Ruixiang Tang 외

Existing bias mitigation methods for DNN models primarily work on learning debiased encoders. This process not only requires a lot of instance-level annotations for sensitive attributes, it also does not guarantee that a…

AttributeClassificationFairness

FairCLIP: Social Bias Elimination based on Attribute Prototype Learning and Representation Neutralization

2022-10-26 · Junyang Wang, Yi Zhang, Jitao Sang

The Vision-Language Pre-training (VLP) models like CLIP have gained popularity in recent years. However, many works found that the social biases hidden in CLIP easily manifest in downstream tasks, especially in image ret…

AttributeFairnessImage RetrievalRetrieval

FairSIN: Achieving Fairness in Graph Neural Networks through Sensitive Information Neutralization

2024-03-19 · Cheng Yang, Jixi Liu, Yunhe Yan, Chuan Shi

Despite the remarkable success of graph neural networks (GNNs) in modeling graph-structured data, like other machine learning models, GNNs are also susceptible to making biased predictions based on sensitive attributes, …

Fairness

Measuring Bias in a Ranked List using Term-based Representations

2024-03-09 · Amin Abolghasemi, Leif Azzopardi, Arian Askari, Maarten de Rijke 외

In most recent studies, gender bias in document ranking is evaluated with the NFaiRR metric, which measures bias in a ranked list based on an aggregation over the unbiasedness scores of each ranked document. This perspec…

Document RankingFairnessPassage Ranking

Prompt Fairness: Sub-group Disparities in LLMs

2025-11-25 · Meiyu Zhong, Noel Teku, Ravi Tandon arxiv

Large Language Models (LLMs), though shown to be effective in many applications, can vary significantly in their response quality. In this paper, we investigate this problem of prompt fairness: specifically, the phrasing…