paper-with-me

홈 › Papers

Impact of Gender Debiased Word Embeddings in Language Modeling

2021-05-03 · Christine Basta, Marta R. Costa-jussà

Gender, race and social biases have recently been detected as evident examples of unfairness in applications of Natural Language Processing. A key path towards fairness is to understand, analyse and interpret our data and algorithms. Recent studies have shown that the human-generated data used in training is an apparent factor of getting biases. In addition, current algorithms have also been proven to amplify biases from data. To further address these concerns, in this paper, we study how an state-of-the-art recurrent neural language model behaves when trained on data, which under-represents females, using pre-trained standard and debiased word embeddings. Results show that language models inherit higher bias when trained on unbalanced data when using pre-trained embeddings, in comparison with using embeddings trained within the task. Moreover, results show that, on the same data, language models inherit lower bias when using debiased pre-trained emdeddings, compared to using standard pre-trained embeddings.

📄 PDF Abstract BibTeX arXiv:2105.00908

Code (0)

등록된 구현이 없습니다.

Tasks

FairnessLanguage ModelingLanguage ModellingWord Embeddings

Similar Papers 제목 키워드 기반

Evaluating the Underlying Gender Bias in Contextualized Word Embeddings

2019-04-18 · WS 2019 8 · Christine Basta, Marta R. Costa-jussà, Noe Casas

Gender bias is highly impacting natural language processing applications. Word embeddings have clearly been proven both to keep and amplify gender biases that are present in current data sources. Recently, contextualized…

SentenceWord Embeddings

Evaluating Bias In Dutch Word Embeddings

2020-10-31 · GeBNLP (COLING) 2020 12 · Rodrigo Alejandro Chávez Mulsa, Gerasimos Spanakis

Recent research in Natural Language Processing has revealed that word embeddings can encode social biases present in the training data which can affect minorities in real world applications. This paper explores the gende…

ClusteringSentenceSentence EmbeddingsWord Embeddings

Reducing Gender Bias in Abusive Language Detection

2018-08-22 · EMNLP 2018 10 · Ji Ho Park, Jamin Shin, Pascale Fung

Abusive language detection models tend to have a problem of being biased toward identity words of a certain group of people because of imbalanced training datasets. For example, "You are a good woman" was considered "sex…

Abusive LanguageData AugmentationWord Embeddings

Unsupervised Mitigating Gender Bias by Character Components: A Case Study of Chinese Word Embedding

2022-07-01 · NAACL (GeBNLP) 2022 7 · Xiuying Chen, Mingzhe Li, Rui Yan, Xin Gao 외

Word embeddings learned from massive text collections have demonstrated significant levels of discriminative biases.However, debias on the Chinese language, one of the most spoken languages, has been less explored.Meanwh…

Word Embeddings

Lipstick on a Pig: Debiasing Methods Cover up Systematic Gender Biases in Word Embeddings But do not Remove Them

2019-03-09 · NAACL 2019 6 · Hila Gonen, Yoav Goldberg

Word embeddings are widely used in NLP for a vast range of tasks. It was shown that word embeddings derived from text corpora reflect gender biases in society. This phenomenon is pervasive and consistent across different…

Word Embeddings