paper-with-me

홈 › Papers

Investigating Annotator Bias in Abusive Language Datasets

2021-09-01 · RANLP 2021 9 · Maximilian Wich, Christian Widmer, Gerhard Hagerer, Georg Groh

Nowadays, social media platforms use classification models to cope with hate speech and abusive language. The problem of these models is their vulnerability to bias. A prevalent form of bias in hate speech and abusive language datasets is annotator bias caused by the annotator’s subjective perception and the complexity of the annotation task. In our paper, we develop a set of methods to measure annotator bias in abusive language datasets and to identify different perspectives on abusive language. We apply these methods to four different abusive language datasets. Our proposed approach supports annotation processes of such datasets and future research addressing different perspectives on the perception of abusive language.

📄 PDF Abstract BibTeX

Code (1)

mawic/annotator-bias-abusive-language 공식 구현

Tasks

Abusive Language

Similar Papers 제목 키워드 기반

Investigating Sampling Bias in Abusive Language Detection

2020-11-01 · EMNLP (ALW) 2020 11 · Dante Razo, Sandra Kübler

Abusive language detection is becoming increasingly important, but we still understand little about the biases in our datasets for abusive language detection, and how these biases affect the quality of abusive language d…

Abusive Language

Lower Bias, Higher Density Abusive Language Datasets: A Recipe

2020-05-01 · LREC 2020 5 · Juliet van Rosendaal, Tommaso Caselli, Malvina Nissim

Datasets to train models for abusive language detection are at the same time necessary and still scarce. One the reasons for their limited availability is the cost of their creation. It is not only that manual annotation…

Abusive Language

Identifying and Measuring Annotator Bias Based on Annotators’ Demographic Characteristics

2020-11-01 · EMNLP (ALW) 2020 11 · Hala Al Kuwatly, Maximilian Wich, Georg Groh

Machine learning is recently used to detect hate speech and other forms of abusive language in online platforms. However, a notable weakness of machine learning models is their vulnerability to bias, which can impair the…

Abusive LanguageBIG-bench Machine LearningFairness

Racial Bias in Hate Speech and Abusive Language Detection Datasets

2019-05-29 · WS 2019 8 · Thomas Davidson, Debasmita Bhattacharya, Ingmar Weber

Technologies for abusive language detection are being developed and applied with little consideration of their potential biases. We examine racial bias in five different sets of Twitter data annotated for hate speech and…

Abuse DetectionAbusive Language

Toward Annotator Group Bias in Crowdsourcing

2021-10-08 · ACL 2022 5 · Haochen Liu, Joseph Thekinen, Sinem Mollaoglu, Da Tang 외

Crowdsourcing has emerged as a popular approach for collecting annotated data to train supervised machine learning models. However, annotator bias can lead to defective annotations. Though there are a few works investiga…