paper-with-me

홈 › Papers

Algorithmic Fairness in NLP: Persona-Infused LLMs for Human-Centric Hate Speech Detection

2025-10-22 · Ewelina Gajewska, Arda Derbent, Jaroslaw A Chudziak, Katarzyna Budzynska arxiv

In this paper, we investigate how personalising Large Language Models (Persona-LLMs) with annotator personas affects their sensitivity to hate speech, particularly regarding biases linked to shared or differing identities between annotators and targets. To this end, we employ Google's Gemini and OpenAI's GPT-4.1-mini models and two persona-prompting methods: shallow persona prompting and a deeply contextualised persona development based on Retrieval-Augmented Generation (RAG) to incorporate richer persona profiles. We analyse the impact of using in-group and out-group annotator personas on the models' detection performance and fairness across diverse social groups. This work bridges psychological insights on group identity with advanced NLP techniques, demonstrating that incorporating socio-demographic attributes into LLMs can address bias in automated hate speech detection. Our results highlight both the potential and limitations of persona-based approaches in reducing bias, offering valuable insights for developing more equitable hate speech detection systems.

📄 PDF Abstract BibTeX arXiv:2510.19331

Code (0)

등록된 구현이 없습니다.

Tasks

Hate Speech Detection

Similar Papers 제목 키워드 기반

Dimensions of Diversity in Human Perceptions of Algorithmic Fairness

2020-05-02 · Nina Grgić-Hlača, Gabriel Lima, Adrian Weller, Elissa M. Redmiles

A growing number of oversight boards and regulatory bodies seek to monitor and govern algorithms that make decisions about people's lives. Prior work has explored how people believe algorithmic decisions should be made, …

Decision MakingDiversityFairness

PsyPlay: Personality-Infused Role-Playing Conversational Agents

2025-02-06 · Tao Yang, Yuhua Zhu, Xiaojun Quan, Cong Liu 외

The current research on Role-Playing Conversational Agents (RPCAs) with Large Language Models (LLMs) primarily focuses on imitating specific speaking styles and utilizing character backgrounds, neglecting the depiction o…

Dialogue Generation

Dynamic fairness-aware recommendation through multi-agent social choice

2023-03-02 · Amanda Aird, Paresha Farastu, Joshua Sun, Elena Štefancová 외

Algorithmic fairness in the context of personalized recommendation presents significantly different challenges to those commonly encountered in classification tasks. Researchers studying classification have generally con…

FairnessRecommendation Systems

The FairCeptron: A Framework for Measuring Human Perceptions of Algorithmic Fairness

2021-02-08 · Georg Ahnert, Ivan Smirnov, Florian Lemmerich, Claudia Wagner 외

Measures of algorithmic fairness often do not account for human perceptions of fairness that can substantially vary between different sociodemographics and stakeholders. The FairCeptron framework is an approach for study…

Decision MakingFairness

The Better Angels of Machine Personality: How Personality Relates to LLM Safety

2024-07-17 · Jie Zhang, Dongrui Liu, Chen Qian, Ziyue Gan 외

Personality psychologists have analyzed the relationship between personality and safety behaviors in human society. Although Large Language Models (LLMs) demonstrate personality traits, the relationship between personali…

FairnessSafety Alignment