paper-with-me

Papers

Annotators with Attitudes: How Annotator Beliefs And Identities Bias Toxic Language Detection

2021-11-15 · NAACL 2022 7 · Maarten Sap, Swabha Swayamdipta, Laura Vianna, Xuhui Zhou, Yejin Choi, Noah A. Smith

The perceived toxicity of language can vary based on someone's identity and beliefs, but this variation is often ignored when collecting toxic language datasets, resulting in dataset and model biases. We seek to understand the who, why, and what behind biases in toxicity annotations. In two online studies with demographically and politically diverse participants, we investigate the effect of annotator identities (who) and beliefs (why), drawing from social psychology research about hate speech, free speech, racist beliefs, political leaning, and more. We disentangle what is annotated as toxic by considering posts with three characteristics: anti-Black language, African American English (AAE) dialect, and vulgarity. Our results show strong associations between annotator identity and beliefs and their ratings of toxicity. Notably, more conservative annotators and those who scored highly on our scale for racist beliefs were less likely to rate anti-Black language as toxic, but more likely to rate AAE as toxic. We additionally present a case study illustrating how a popular toxicity detection system's ratings inherently reflect only specific beliefs and perspectives. Our findings call for contextualizing toxicity labels in social variables, which raises immense implications for toxic language annotation and detection.

📄 PDF Abstract BibTeX arXiv:2111.07997

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

American 설명 없음

Similar Papers 제목 키워드 기반

Reducing annotator bias by belief elicitation

2024-10-21 · Terne Sasha Thorn Jakobsen, Andreas Bjerre-Nielsen, Robert Böhm

Crowdsourced annotations of data play a substantial role in the development of Artificial Intelligence (AI). It is broadly recognised that annotations of text data can contain annotator bias, where systematic disagreemen…

Re-examining Sexism and Misogyny Classification with Annotator Attitudes

2024-10-04 · Aiqi Jiang, Nikolas Vitsakis, Tanvi Dinkar, Gavin Abercrombie 외

Gender-Based Violence (GBV) is an increasing problem online, but existing datasets fail to capture the plurality of possible annotator perspectives or ensure the representation of affected groups. We revisit two importan…

When Annotators Agree but Labels Disagree: The Projection Problem in Stance Detection

2026-03-25 · Bowen Zhang arxiv

Stance detection is nearly always formulated as classifying text into Favor, Against, or Neutral. This convention was inherited from debate analysis and has been applied without modification to social media since SemEval…

Stance Detection

MBIC -- A Media Bias Annotation Dataset Including Annotator Characteristics

2021-05-20 · T. Spinde, L. Rudnitckaia, K. Sinha, F. Hamborg 외

Many people consider news articles to be a reliable source of information on current events. However, due to the range of factors influencing news agencies, such coverage may not always be impartial. Media bias, or slant…

ArticlesBias DetectionSentence

Towards Equal Gender Representation in the Annotations of Toxic Language Detection

2021-06-04 · ACL (GeBNLP) 2021 8 · Elizabeth Excell, Noura Al Moubayed

Classifiers tend to propagate biases present in the data on which they are trained. Hence, it is important to understand how the demographic identities of the annotators of comments affect the fairness of the resulting m…

Fairness