paper-with-me

홈 › Papers

BiasTestGPT: Using ChatGPT for Social Bias Testing of Language Models

2023-02-14 · Rafal Kocielnik, Shrimai Prabhumoye, Vivian Zhang, Roy Jiang, R. Michael Alvarez, Anima Anandkumar

Pretrained Language Models (PLMs) harbor inherent social biases that can result in harmful real-world implications. Such social biases are measured through the probability values that PLMs output for different social groups and attributes appearing in a set of test sentences. However, bias testing is currently cumbersome since the test sentences are generated either from a limited set of manual templates or need expensive crowd-sourcing. We instead propose using ChatGPT for the controllable generation of test sentences, given any arbitrary user-specified combination of social groups and attributes appearing in the test sentences. When compared to template-based methods, our approach using ChatGPT for test sentence generation is superior in detecting social bias, especially in challenging settings such as intersectional biases. We present an open-source comprehensive bias testing framework (BiasTestGPT), hosted on HuggingFace, that can be plugged into any open-source PLM for bias testing. User testing with domain experts from various fields has shown their interest in being able to test modern AI for social biases. Our tool has significantly improved their awareness of such biases in PLMs, proving to be learnable and user-friendly. We thus enable seamless open-ended social bias testing of PLMs by domain experts through an automatic large-scale generation of diverse test sentences for any combination of social categories and attributes.

📄 PDF Abstract BibTeX arXiv:2302.07371

Code (0)

등록된 구현이 없습니다.

Tasks

SentenceText Generation

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

Testing and Evaluation of Large Language Models: Correctness, Non-Toxicity, and Fairness

2024-08-31 · Wenxuan Wang

Large language models (LLMs), such as ChatGPT, have rapidly penetrated into people's work and daily lives over the past few years, due to their extraordinary conversational skills and intelligence. ChatGPT has become the…

FairnessLanguage ModelingLanguage ModellingLogical Reasoning+2

Pipelines for Social Bias Testing of Large Language Models

2022-05-01 · BigScience (ACL) 2022 5 · Debora Nozza, Federico Bianchi, Dirk Hovy

The maturity level of language models is now at a stage in which many companies rely on them to solve various tasks. However, while research has shown how biased and harmful these models are, systematic ways of integrati…

software testing

Large Language Models Portray Socially Subordinate Groups as More Homogeneous, Consistent with a Bias Observed in Humans

2024-01-16 · Messi H. J. Lee, Jacob M. Montgomery, Calvin K. Lai

Large language models (LLMs) are becoming pervasive in everyday life, yet their propensity to reproduce biases inherited from training data remains a pressing concern. Prior investigations into bias in LLMs have focused …

Public Perceptions of Gender Bias in Large Language Models: Cases of ChatGPT and Ernie

2023-09-17 · Kyrie Zhixuan Zhou, Madelyn Rose Sanfilippo

Large language models are quickly gaining momentum, yet are found to demonstrate gender bias in their responses. In this paper, we conducted a content analysis of social media discussions to gauge public perceptions of g…

I Am Not Them: Fluid Identities and Persistent Out-group Bias in Large Language Models

2024-02-16 · Wenchao Dong, Assem Zhunis, Hyojin Chin, Jiyoung Han 외

We explored cultural biases-individualism vs. collectivism-in ChatGPT across three Western languages (i.e., English, German, and French) and three Eastern languages (i.e., Chinese, Japanese, and Korean). When ChatGPT ado…

Prompt Engineering