paper-with-me

홈 › Papers

Mitigating Biases to Embrace Diversity: A Comprehensive Annotation Benchmark for Toxic Language

2024-10-17 · Xinmeng Hou

This study introduces a prescriptive annotation benchmark grounded in humanities research to ensure consistent, unbiased labeling of offensive language, particularly for casual and non-mainstream language uses. We contribute two newly annotated datasets that achieve higher inter-annotator agreement between human and language model (LLM) annotations compared to original datasets based on descriptive instructions. Our experiments show that LLMs can serve as effective alternatives when professional annotators are unavailable. Moreover, smaller models fine-tuned on multi-source LLM-annotated data outperform models trained on larger, single-source human-annotated datasets. These findings highlight the value of structured guidelines in reducing subjective variability, maintaining performance with limited data, and embracing language diversity. Content Warning: This article only analyzes offensive language for academic purposes. Discretion is advised.

📄 PDF Abstract BibTeX arXiv:2410.13313

Code (0)

등록된 구현이 없습니다.

Tasks

DescriptiveDiversityLanguage ModelingLanguage Modelling

Similar Papers 제목 키워드 기반

Mitigating Bias in Facial Analysis Systems by Incorporating Label Diversity

2022-04-13 · Camila Kolling, Victor Araujo, Adriano Veloso, Soraia Raupp Musse

Facial analysis models are increasingly applied in real-world applications that have significant impact on peoples' lives. However, as literature has shown, models that automatically classify facial attributes might exhi…

DiversityEnsemble Learning

Evaluating Gender Bias in Speech Translation

2020-10-27 · LREC 2022 6 · Marta R. Costa-jussà, Christine Basta, Gerard I. Gállego

The scientific community is increasingly aware of the necessity to embrace pluralism and consistently represent major and minor social groups. Currently, there are no standard evaluation techniques for different types of…

Translation

Wisdom from Diversity: Bias Mitigation Through Hybrid Human-LLM Crowds

2025-05-18 · Axel Abels, Tom Lenaerts

Despite their performance, large language models (LLMs) can inadvertently perpetuate biases found in the data they are trained on. By analyzing LLM responses to bias-eliciting headlines, we find that these models often m…

Diversity

Debiased Contrastive Representation Learning for Mitigating Dual Biases in Recommender Systems

2024-08-19 · Zhirong Huang, Shichao Zhang, Debo Cheng, Jiuyong Li 외

In recommender systems, popularity and conformity biases undermine recommender effectiveness by disproportionately favouring popular items, leading to their over-representation in recommendation lists and causing an unba…

Contrastive LearningDiversityRecommendation SystemsRepresentation Learning

Understanding and Mitigating Annotation Bias in Facial Expression Recognition

2021-08-19 · ICCV 2021 10 · Yunliang Chen, Jungseock Joo

The performance of a computer vision model depends on the size and quality of its training data. Recent studies have unveiled previously-unknown composition biases in common image datasets which then lead to skewed model…

Facial Expression RecognitionFacial Expression Recognition (FER)Triplet