paper-with-me

Papers

Evaluating Gender, Racial, and Age Biases in Large Language Models: A Comparative Analysis of Occupational and Crime Scenarios

2024-09-22 · Vishal Mirza, Rahul Kulkarni, Aakanksha Jadhav

Recent advancements in Large Language Models(LLMs) have been notable, yet widespread enterprise adoption remains limited due to various constraints. This paper examines bias in LLMs-a crucial issue affecting their usability, reliability, and fairness. Researchers are developing strategies to mitigate bias, including debiasing layers, specialized reference datasets like Winogender and Winobias, and reinforcement learning with human feedback (RLHF). These techniques have been integrated into the latest LLMs. Our study evaluates gender bias in occupational scenarios and gender, age, and racial bias in crime scenarios across four leading LLMs released in 2024: Gemini 1.5 Pro, Llama 3 70B, Claude 3 Opus, and GPT-4o. Findings reveal that LLMs often depict female characters more frequently than male ones in various occupations, showing a 37% deviation from US BLS data. In crime scenarios, deviations from US FBI data are 54% for gender, 28% for race, and 17% for age. We observe that efforts to reduce gender and racial bias often lead to outcomes that may over-index one sub-class, potentially exacerbating the issue. These results highlight the limitations of current bias mitigation techniques and underscore the need for more effective approaches.

📄 PDF Abstract BibTeX arXiv:2409.14583

Code (0)

등록된 구현이 없습니다.

Tasks

Fairness

Methods 이 논문이 사용한 방법론

LLaMA LLaMA is a collection of foundation language models ranging from 7B to 65B parameters. It is based on the transformer architecture with various improvements that were…

Similar Papers 제목 키워드 기반

Evaluating Large Language Models through Gender and Racial Stereotypes

2023-11-24 · Ananya Malik

Language Models have ushered a new age of AI gaining traction within the NLP community as well as amongst the general population. AI's ability to make predictions, generations and its applications in sensitive decision-m…

Decision Making

Evaluation of Bias Towards Medical Professionals in Large Language Models

2024-06-30 · Xi Chen, Yang Xu, MingKe You, Li Wang 외

This study evaluates whether large language models (LLMs) exhibit biases towards medical professionals. Fictitious candidate resumes were created to control for identity factors while maintaining consistent qualification…

ChatGPT Exhibits Gender and Racial Biases in Acute Coronary Syndrome Management

2023-11-10 · Angela Zhang, Mert Yuksekgonul, Joshua Guild, James Zou 외

Recent breakthroughs in large language models (LLMs) have led to their rapid dissemination and widespread use. One early application has been to medicine, where LLMs have been investigated to streamline clinical workflow…

Decision MakingManagement

Gender and Racial Stereotype Detection in Legal Opinion Word Embeddings

2022-03-24 · Sean Matthews, John Hudzina, Dawn Sepehr

Studies have shown that some Natural Language Processing (NLP) systems encode and replicate harmful biases with potential adverse ethical effects in our society. In this article, we propose an approach for identifying ge…

Question AnsweringWord Embeddings

Understanding and Evaluating Racial Biases in Image Captioning

2021-06-16 · ICCV 2021 10 · Dora Zhao, Angelina Wang, Olga Russakovsky

Image captioning is an important task for benchmarking visual reasoning and for enabling accessibility for people with vision impairments. However, as in many machine learning settings, social biases can influence image …

BenchmarkingImage CaptioningVisual Reasoning