paper-with-me

홈 › Papers

Benchmarking Intersectional Biases in NLP

2022-07-01 · NAACL 2022 7 · John Lalor, Yi Yang, Kendall Smith, Nicole Forsgren, Ahmed Abbasi

There has been a recent wave of work assessing the fairness of machine learning models in general, and more specifically, on natural language processing (NLP) models built using machine learning techniques. While much work has highlighted biases embedded in state-of-the-art language models, and more recent efforts have focused on how to debias, research assessing the fairness and performance of biased/debiased models on downstream prediction tasks has been limited. Moreover, most prior work has emphasized bias along a single dimension such as gender or race. In this work, we benchmark multiple NLP models with regards to their fairness and predictive performance across a variety of NLP tasks. In particular, we assess intersectional bias - fairness across multiple demographic dimensions. The results show that while current debiasing strategies fare well in terms of the fairness-accuracy trade-off (generally preserving predictive power in debiased models), they are unable to effectively alleviate bias in downstream tasks. Furthermore, this bias is often amplified across dimensions (i.e., intersections). We conclude by highlighting possible causes and making recommendations for future NLP debiasing research.

📄 PDF Abstract BibTeX

Code (1)

nd-hal/naacl-2022 공식 구현

Tasks

BenchmarkingBIG-bench Machine LearningFairness

Similar Papers 제목 키워드 기반

IndiBias: A Benchmark Dataset to Measure Social Biases in Language Models for Indian Context

2024-03-29 · Nihar Ranjan Sahoo, Pranamya Prashant Kulkarni, Narjis Asad, Arif Ahmad 외

The pervasive influence of social biases in language data has sparked the need for benchmark datasets that capture and evaluate these biases in Large Language Models (LLMs). Existing efforts predominantly focus on Englis…

BenchmarkingSentence

Detecting Emergent Intersectional Biases: Contextualized Word Embeddings Contain a Distribution of Human-like Biases

2020-06-06 · Wei Guo, Aylin Caliskan

With the starting point that implicit human biases are reflected in the statistical regularities of language, it is possible to measure biases in English static word embeddings. State-of-the-art neural language models ge…

Bias DetectionSentenceWord Embeddings

WordBias: An Interactive Visual Tool for Discovering Intersectional Biases Encoded in Word Embeddings

2021-03-05 · Bhavya Ghai, Md Naimul Hoque, Klaus Mueller

Intersectional bias is a bias caused by an overlap of multiple social factors like gender, sexuality, race, disability, religion, etc. A recent study has shown that word embedding models can be laden with biases against …

Word Embeddings

Mapping the Multilingual Margins: Intersectional Biases of Sentiment Analysis Systems in English, Spanish, and Arabic

2022-04-07 · LTEDI (ACL) 2022 5 · António Câmara, Nina Taneja, Tamjeed Azad, Emily Allaway 외

As natural language processing systems become more widespread, it is necessary to address fairness issues in their implementation and deployment to ensure that their negative impacts on society are understood and minimiz…

FairnessregressionSentiment Analysis

White Men Lead, Black Women Help? Benchmarking and Mitigating Language Agency Social Biases in LLMs

2024-04-16 · Yixin Wan, Kai-Wei Chang

Social biases can manifest in language agency. However, very limited research has investigated such biases in Large Language Model (LLM)-generated content. In addition, previous works often rely on string-matching techni…

BenchmarkingLanguage ModellingLarge Language ModelSentence+1