paper-with-me

Papers

Gender and content bias in Large Language Models: a case study on Google Gemini 2.0 Flash Experimental

2025-03-18 · Roberto Balestri

This study evaluates the biases in Gemini 2.0 Flash Experimental, a state-of-the-art large language model (LLM) developed by Google, focusing on content moderation and gender disparities. By comparing its performance to ChatGPT-4o, examined in a previous work of the author, the analysis highlights some differences in ethical moderation practices. Gemini 2.0 demonstrates reduced gender bias, notably with female-specific prompts achieving a substantial rise in acceptance rates compared to results obtained by ChatGPT-4o. It adopts a more permissive stance toward sexual content and maintains relatively high acceptance rates for violent prompts, including gender-specific cases. Despite these changes, whether they constitute an improvement is debatable. While gender bias has been reduced, this reduction comes at the cost of permitting more violent content toward both males and females, potentially normalizing violence rather than mitigating harm. Male-specific prompts still generally receive higher acceptance rates than female-specific ones. These findings underscore the complexities of aligning AI systems with ethical standards, highlighting progress in reducing certain biases while raising concerns about the broader implications of the model's permissiveness. Ongoing refinements are essential to achieve moderation practices that ensure transparency, fairness, and inclusivity without amplifying harmful content.

📄 PDF Abstract BibTeX arXiv:2503.16534

Code (0)

등록된 구현이 없습니다.

Tasks

FairnessLarge Language Model

Similar Papers 제목 키워드 기반

Public Perceptions of Gender Bias in Large Language Models: Cases of ChatGPT and Ernie

2023-09-17 · Kyrie Zhixuan Zhou, Madelyn Rose Sanfilippo

Large language models are quickly gaining momentum, yet are found to demonstrate gender bias in their responses. In this paper, we conducted a content analysis of social media discussions to gauge public perceptions of g…

LLMs Reproduce Stereotypes of Sexual and Gender Minorities

2025-01-10 · Ruby Ostrow, Adam Lopez

A large body of research has found substantial gender bias in NLP systems. Most of this research takes a binary, essentialist view of gender: limiting its variation to the categories _men_ and _women_, conflating gender …

Text Generation

Bias of AI-Generated Content: An Examination of News Produced by Large Language Models

2023-09-18 · Xiao Fang, Shangkun Che, Minjia Mao, Hongzhe Zhang 외

Large language models (LLMs) have the potential to transform our lives and work through the content they generate, known as AI-Generated Content (AIGC). To harness this transformation, we need to understand the limitatio…

Articles

How Gender Interacts with Political Values: A Case Study on Czech BERT Models

2024-03-20 · Adnan Al Ali, Jindřich Libovický

Neural language models, which reach state-of-the-art results on most natural language processing tasks, are trained on large text corpora that inevitably contain value-burdened content and often capture undesirable biase…

Survey

GenderAlign: An Alignment Dataset for Mitigating Gender Bias in Large Language Models

2024-06-20 · Tao Zhang, Ziqian Zeng, Yuxiang Xiao, Huiping Zhuang 외

Large Language Models (LLMs) are prone to generating content that exhibits gender biases, raising significant ethical concerns. Alignment, the process of fine-tuning LLMs to better align with desired behaviors, is recogn…

8k