paper-with-me

홈 › Papers

Challenges in Annotating Datasets to Quantify Bias in Under-represented Society

2023-09-11 · Vithya Yogarajan, Gillian Dobbie, Timothy Pistotti, Joshua Bensemann, Kobe Knowles

Recent advances in artificial intelligence, including the development of highly sophisticated large language models (LLM), have proven beneficial in many real-world applications. However, evidence of inherent bias encoded in these LLMs has raised concerns about equity. In response, there has been an increase in research dealing with bias, including studies focusing on quantifying bias and developing debiasing techniques. Benchmark bias datasets have also been developed for binary gender classification and ethical/racial considerations, focusing predominantly on American demographics. However, there is minimal research in understanding and quantifying bias related to under-represented societies. Motivated by the lack of annotated datasets for quantifying bias in under-represented societies, we endeavoured to create benchmark datasets for the New Zealand (NZ) population. We faced many challenges in this process, despite the availability of three annotators. This research outlines the manual annotation process, provides an overview of the challenges we encountered and lessons learnt, and presents recommendations for future research.

📄 PDF Abstract BibTeX arXiv:2309.08624

Code (0)

등록된 구현이 없습니다.

Tasks

Gender Classification

Methods 이 논문이 사용한 방법론

American 설명 없음

Similar Papers 제목 키워드 기반

Is Your Classifier Actually Biased? Measuring Fairness under Uncertainty with Bernstein Bounds

2020-04-26 · ACL 2020 6 · Kawin Ethayarajh

Most NLP datasets are not annotated with protected attributes such as gender, making it difficult to measure classification bias using standard measures of fairness (e.g., equal opportunity). However, manually annotating…

AttributeFairness

Quantifying Political Bias in News Articles

2022-10-07 · Gizem Gezici

Search bias analysis is getting more attention in recent years since search results could affect In this work, we aim to establish an automated model for evaluating ideological bias in online news articles. The dataset i…

Articles

Challenges in Measuring Bias via Open-Ended Language Generation

2022-05-23 · NAACL (GeBNLP) 2022 7 · Afra Feyza Akyürek, Muhammed Yusuf Kocyigit, Sejin Paik, Derry Wijaya

Researchers have devised numerous ways to quantify social biases vested in pretrained language models. As some language models are capable of generating coherent completions given a set of textual prompts, several prompt…

Language ModelingLanguage ModellingText Generation

Annotating Online Misogyny

2021-08-01 · ACL 2021 5 · Philine Zeinert, Nanna Inie, Leon Derczynski

Online misogyny, a category of online abusive language, has serious and harmful social consequences. Automatic detection of misogynistic language online, while imperative, poses complicated challenges to both data gather…

Abusive LanguageHate Speech Detection

GeoBS: Information-Theoretic Quantification of Geographic Bias in AI Models

2025-09-27 · Zhangyu Wang, Nemin Wu, Qian Cao, Jiangnan Xia 외 arxiv

The widespread adoption of AI models, especially foundation models (FMs), has made a profound impact on numerous domains. However, it also raises significant ethical concerns, including bias issues. Although numerous eff…