paper-with-me

Papers

Large Language Models Still Exhibit Bias in Long Text

2024-10-23 · Wonje Jeung, Dongjae Jeon, Ashkan Yousefpour, Jonghyun Choi

Existing fairness benchmarks for large language models (LLMs) primarily focus on simple tasks, such as multiple-choice questions, overlooking biases that may arise in more complex scenarios like long-text generation. To address this gap, we introduce the Long Text Fairness Test (LTF-TEST), a framework that evaluates biases in LLMs through essay-style prompts. LTF-TEST covers 14 topics and 10 demographic axes, including gender and race, resulting in 11,948 samples. By assessing both model responses and the reasoning behind them, LTF-TEST uncovers subtle biases that are difficult to detect in simple responses. In our evaluation of five recent LLMs, including GPT-4o and LLaMa3, we identify two key patterns of bias. First, these models frequently favor certain demographic groups in their responses. Second, they show excessive sensitivity toward traditionally disadvantaged groups, often providing overly protective responses while neglecting others. To mitigate these biases, we propose FT-REGARD, a finetuning approach that pairs biased prompts with neutral responses. FT-REGARD reduces gender bias by 34.6% and improves performance by 1.4 percentage points on the BBQ benchmark, offering a promising approach to addressing biases in long-text generation tasks.

📄 PDF Abstract BibTeX arXiv:2410.17519

Code (0)

등록된 구현이 없습니다.

Tasks

FairnessMultiple-choiceText Generation

Methods 이 논문이 사용한 방법론

Focus 설명 없음

Similar Papers 제목 키워드 기반

Spoken Stereoset: On Evaluating Social Bias Toward Speaker in Speech Large Language Models

2024-08-14 · Yi-Cheng Lin, Wei-Chih Chen, Hung-Yi Lee

Warning: This paper may contain texts with uncomfortable content. Large Language Models (LLMs) have achieved remarkable performance in various tasks, including those involving multimodal data like speech. However, these …

Quantifying Gender Bias in Large Language Models: When ChatGPT Becomes a Hiring Manager

2026-03-10 · Nina Gerszberg, Janka Hamori, Andrew Lo arxiv

The growing prominence of large language models (LLMs) in daily life has heightened concerns that LLMs exhibit many of the same gender-related biases as their creators. In the context of hiring decisions, we quantify the…

Prompt Engineering

Actions Speak Louder than Words: Agent Decisions Reveal Implicit Biases in Language Models

2025-01-29 · YuXuan Li, Hirokazu Shirado, Sauvik Das

While advances in fairness and alignment have helped mitigate overt biases exhibited by large language models (LLMs) when explicitly prompted, we hypothesize that these models may still exhibit implicit biases when simul…

Decision MakingFairness

RobBERTje: a Distilled Dutch BERT Model

2022-04-28 · Pieter Delobelle, Thomas Winters, Bettina Berendt

Pre-trained large-scale language models such as BERT have gained a lot of attention thanks to their outstanding performance on a wide range of natural language tasks. However, due to their large number of parameters, the…

Lightweight Deploymentmodel

"As Eastern Powers, I will veto." : An Investigation of Nation-level Bias of Large Language Models in International Relations

2025-11-12 · Jonghyeon Choi, Yeonjun Choi, Hyun-chul Kim, Beakcheol Jang arxiv

This paper systematically examines nation-level biases exhibited by Large Language Models (LLMs) within the domain of International Relations (IR). Leveraging historical records from the United Nations Security Council (…