paper-with-me

홈 › Papers

The Tail Wagging the Dog: Dataset Construction Biases of Social Bias Benchmarks

2022-10-18 · Nikil Roashan Selvam, Sunipa Dev, Daniel Khashabi, Tushar Khot, Kai-Wei Chang

How reliably can we trust the scores obtained from social bias benchmarks as faithful indicators of problematic social biases in a given language model? In this work, we study this question by contrasting social biases with non-social biases stemming from choices made during dataset construction that might not even be discernible to the human eye. To do so, we empirically simulate various alternative constructions for a given benchmark based on innocuous modifications (such as paraphrasing or random-sampling) that maintain the essence of their social bias. On two well-known social bias benchmarks (Winogender and BiasNLI) we observe that these shallow modifications have a surprising effect on the resulting degree of bias across various models. We hope these troubling observations motivate more robust measures of social biases.

📄 PDF Abstract BibTeX arXiv:2210.10040

Code (1)

uclanlp/socialbias-dataset-construction-biases 공식 구현

Tasks

Language ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

JUBAKU: An Adversarial Benchmark for Exposing Culturally Grounded Stereotypes in Japanese LLMs

2026-03-21 · Taihei Shiotani, Masahiro Kaneko, Ayana Niwa, Yuki Maruyama 외 arxiv

Social biases reflected in language are inherently shaped by cultural norms, which vary significantly across regions and lead to diverse manifestations of stereotypes. Existing evaluations of social bias in large languag…

BanStereoSet: A Dataset to Measure Stereotypical Social Biases in LLMs for Bangla

2024-09-18 · Mahammed Kamruzzaman, Abdullah Al Monsur, Shrabon Das, Enamul Hassan 외

This study presents BanStereoSet, a dataset designed to evaluate stereotypical social biases in multilingual LLMs for the Bangla language. In an effort to extend the focus of bias research beyond English-centric datasets…

Measuring Social Biases in Grounded Vision and Language Embeddings

2020-02-20 · NAACL 2021 4 · Candace Ross, Boris Katz, Andrei Barbu

We generalize the notion of social biases from language embeddings to grounded vision and language embeddings. Biases are present in grounded embeddings, and indeed seem to be equally or more significant than for ungroun…

Word Embeddings

Detecting Unintended Social Bias in Toxic Language Datasets

2022-10-21 · Nihar Sahoo, Himanshu Gupta, Pushpak Bhattacharyya

With the rise of online hate speech, automatic detection of Hate Speech, Offensive texts as a natural language processing task is getting popular. However, very little research has been done to detect unintended social b…

MABR: Multilayer Adversarial Bias Removal Without Prior Bias Knowledge

2024-08-10 · Maxwell J. Yin, Boyu Wang, Charles Ling

Models trained on real-world data often mirror and exacerbate existing social biases. Traditional methods for mitigating these biases typically require prior knowledge of the specific biases to be addressed, such as gend…

Attribute