paper-with-me

Papers

Measuring Social Biases in Masked Language Models by Proxy of Prediction Quality

2024-02-21 · Rahul Zalkikar, Kanchan Chandra

Innovative transformer-based language models produce contextually-aware token embeddings and have achieved state-of-the-art performance for a variety of natural language tasks, but have been shown to encode unwanted biases for downstream applications. In this paper, we evaluate the social biases encoded by transformers trained with the masked language modeling objective using proposed proxy functions within an iterative masking experiment to measure the quality of transformer models' predictions, and assess the preference of MLMs towards disadvantaged and advantaged groups. We compare bias estimations with those produced by other evaluation methods using benchmark datasets and assess their alignment with human annotated biases. We find relatively high religious and disability biases across considered MLMs and low gender bias in one dataset relative to another. We extend on previous work by evaluating social biases introduced after retraining an MLM under the masked language modeling objective, and find that proposed measures produce more accurate estimations of biases introduced by retraining MLMs than others based on relative preference for biased sentences between models.

📄 PDF Abstract BibTeX arXiv:2402.13954

Code (1)

zalkikar/mlm-bias 공식 구현 pytorch

Tasks

Language ModelingLanguage ModellingMasked Language Modeling

Similar Papers 제목 키워드 기반

CrowS-Pairs: A Challenge Dataset for Measuring Social Biases in Masked Language Models

2020-09-30 · EMNLP 2020 11 · Nikita Nangia, Clara Vania, Rasika Bhalerao, Samuel R. Bowman

Pretrained language models, especially masked language models (MLMs) have seen success across many NLP tasks. However, there is ample evidence that they use the cultural biases that are undoubtedly present in the corpora…

Evaluating Short-Term Temporal Fluctuations of Social Biases in Social Media Data and Masked Language Models

2024-06-19 · Yi Zhou, Danushka Bollegala, Jose Camacho-Collados

Social biases such as gender or racial biases have been reported in language models (LMs), including Masked Language Models (MLMs). Given that MLMs are continuously trained with increasing amounts of additional data coll…

valid

Speciesist Language and Nonhuman Animal Bias in English Masked Language Models

2022-03-10 · Masashi Takeshita, Rafal Rzepka, Kenji Araki

Various existing studies have analyzed what social biases are inherited by NLP models. These biases may directly or indirectly harm people, therefore previous studies have focused only on human attributes. However, until…

Speciesist Language and Nonhuman Animal Bias in English Masked Language Models

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Various existing studies have analyzed what social biases are inherited by NLP models. These biases may directly or indirectly harm people, therefore previous studies have focused only on human attributes. If the social …

Unmasking the Mask -- Evaluating Social Biases in Masked Language Models

2021-04-15 · Masahiro Kaneko, Danushka Bollegala

Masked Language Models (MLMs) have shown superior performances in numerous downstream NLP tasks when used as text encoders. Unfortunately, MLMs also demonstrate significantly worrying levels of social biases. We show tha…

Selection biasSentence