paper-with-me

홈 › Papers

This Prompt is Measuring <MASK>: Evaluating Bias Evaluation in Language Models

2023-05-22 · Seraphina Goldfarb-Tarrant, Eddie Ungless, Esma Balkir, Su Lin Blodgett

Bias research in NLP seeks to analyse models for social biases, thus helping NLP practitioners uncover, measure, and mitigate social harms. We analyse the body of work that uses prompts and templates to assess bias in language models. We draw on a measurement modelling framework to create a taxonomy of attributes that capture what a bias test aims to measure and how that measurement is carried out. By applying this taxonomy to 90 bias tests, we illustrate qualitatively and quantitatively that core aspects of bias test conceptualisations and operationalisations are frequently unstated or ambiguous, carry implicit assumptions, or be mismatched. Our analysis illuminates the scope of possible bias types the field is able to measure, and reveals types that are as yet under-researched. We offer guidance to enable the community to explore a wider section of the possible bias space, and to better close the gap between desired outcomes and experimental design, both for bias and for evaluating language models more broadly.

📄 PDF Abstract BibTeX arXiv:2305.12757

Code (0)

등록된 구현이 없습니다.

Tasks

Experimental Design

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Measuring Social Biases in Masked Language Models by Proxy of Prediction Quality

2024-02-21 · Rahul Zalkikar, Kanchan Chandra

Innovative transformer-based language models produce contextually-aware token embeddings and have achieved state-of-the-art performance for a variety of natural language tasks, but have been shown to encode unwanted bias…

Language ModelingLanguage ModellingMasked Language Modeling

MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation

2026-05-21 · Patryk Bartkowiak, Lennart Petersen, Bartosz Kotrys, Dominik Michels 외 arxiv

Evaluating single-concept personalization in text-to-image diffusion requires measuring both concept preservation, which captures identity fidelity to a reference, and prompt following, which captures whether the generat…

Measuring Bias or Measuring the Task: Understanding the Brittle Nature of LLM Gender Biases

2025-09-04 · Bufan Gao, Elisa Kreiss arxiv

As LLMs are increasingly applied in socially impactful settings, concerns about gender bias have prompted growing efforts both to measure and mitigate such bias. These efforts often rely on evaluation tasks that differ f…

Echoes of Agreement: Argument Driven Opinion Shifts in Large Language Models

2025-08-11 · Avneet Kaur arxiv

There have been numerous studies evaluating bias of LLMs towards political topics. However, how positions towards these topics in model outputs are highly sensitive to the prompt. What happens when the prompt itself is s…

LLMs Are Biased Towards Output Formats! Systematically Evaluating and Mitigating Output Format Bias of LLMs

2024-08-16 · Do Xuan Long, Hai Nguyen Ngoc, Tiviatis Sim, Hieu Dao 외

We present the first systematic evaluation examining format bias in performance of large language models (LLMs). Our approach distinguishes between two categories of an evaluation metric under format constraints to relia…

Instruction FollowingMultiple-choice