paper-with-me

홈 › Papers

Assessing the Reliability of LLMs Annotations in the Context of Demographic Bias and Model Explanation

2025-07-17 · Hadi Mohammadi, Tina Shahedi, Pablo Mosteiro, Massimo Poesio, Ayoub Bagheri, Anastasia Giachanou arxiv

Understanding the sources of variability in annotations is crucial for developing fair NLP systems, especially for tasks like sexism detection where demographic bias is a concern. This study investigates the extent to which annotator demographic features influence labeling decisions compared to text content. Using a Generalized Linear Mixed Model, we quantify this inf luence, finding that while statistically present, demographic factors account for a minor fraction ( 8%) of the observed variance, with tweet content being the dominant factor. We then assess the reliability of Generative AI (GenAI) models as annotators, specifically evaluating if guiding them with demographic personas improves alignment with human judgments. Our results indicate that simplistic persona prompting often fails to enhance, and sometimes degrades, performance compared to baseline models. Furthermore, explainable AI (XAI) techniques reveal that model predictions rely heavily on content-specific tokens related to sexism, rather than correlates of demographic characteristics. We argue that focusing on content-driven explanations and robust annotation protocols offers a more reliable path towards fairness than potentially persona simulation.

📄 PDF Abstract BibTeX arXiv:2507.13138

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Value Portrait: Understanding Values of LLMs with Human-aligned Benchmark

2025-05-02 · Jongwook Han, Dongmin Choi, Woojung Song, Eun-Ju Lee 외

The importance of benchmarks for assessing the values of language models has been pronounced due to the growing need of more authentic, human-aligned responses. However, existing benchmarks rely on human or machine annot…

Which Demographics do LLMs Default to During Annotation?

2024-10-11 · Johannes Schäfer, Aidan Combs, Christopher Bagdon, Jiahui Li 외

Demographics and cultural background of annotators influence the labels they assign in text annotation -- for instance, an elderly woman might find it offensive to read a message addressed to a "bro", but a male teenager…

text annotation

GRASP: A Disagreement Analysis Framework to Assess Group Associations in Perspectives

2023-11-09 · Vinodkumar Prabhakaran, Christopher Homan, Lora Aroyo, Aida Mostafazadeh Davani 외

Human annotation plays a core role in machine learning -- annotations for supervised models, safety guardrails for generative models, and human feedback for reinforcement learning, to cite a few avenues. However, the fac…

Chatbot

Exploring Robustness of LLMs to Sociodemographically-Conditioned Paraphrasing

2025-01-14 · Pulkit Arora, Akbar Karimi, Lucie Flek

Large Language Models (LLMs) have shown impressive performance in various NLP tasks. However, there are concerns about their reliability in different domains of linguistic variations. Many works have proposed robustness …

Comparing LLM Text Annotation Skills: A Study on Human Rights Violations in Social Media Data

2025-05-15 · Poli Apollinaire Nemkova, Solomon Ubani, Mark V. Albert

In the era of increasingly sophisticated natural language processing (NLP) systems, large language models (LLMs) have demonstrated remarkable potential for diverse applications, including tasks requiring nuanced textual …

Binary Classificationtext annotation