paper-with-me

홈 › Papers

A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance

2025-06-23 · Matteo Melis, Gabriella Lapesa, Dennis Assenmacher

Detecting harmful content is a crucial task in the landscape of NLP applications for Social Good, with hate speech being one of its most dangerous forms. But what do we mean by hate speech, how can we define it, and how does prompting different definitions of hate speech affect model performance? The contribution of this work is twofold. At the theoretical level, we address the ambiguity surrounding hate speech by collecting and analyzing existing definitions from the literature. We organize these definitions into a taxonomy of 14 Conceptual Elements-building blocks that capture different aspects of hate speech definitions, such as references to the target of hate (individual or groups) or of the potential consequences of it. At the experimental level, we employ the collection of definitions in a systematic zero-shot evaluation of three LLMs, on three hate speech datasets representing different types of data (synthetic, human-in-the-loop, and real-world). We find that choosing different definitions, i.e., definitions with a different degree of specificity in terms of encoded elements, impacts model performance, but this effect is not consistent across all architectures.

📄 PDF Abstract BibTeX arXiv:2506.18576

Code (1)

matteo-mls/modular-taxonomy-for-hate-speech-definitions 공식 구현

Tasks

Specificity

Similar Papers 제목 키워드 기반

Improving Hate Speech Classification with Cross-Taxonomy Dataset Integration

2025-03-07 · Jan Fillies, Adrian Paschke

Algorithmic hate speech detection faces significant challenges due to the diverse definitions and datasets used in research and practice. Social media platforms, legal frameworks, and institutions each apply distinct yet…

Hate Speech Detection

Untangling Hate Speech Definitions: A Semantic Componential Analysis Across Cultures and Domains

2024-11-11 · Katerina Korre, Arianna Muti, Federico Ruggeri, Alberto Barrón-Cedeño

Hate speech relies heavily on cultural influences, leading to varying individual interpretations. For that reason, we propose a Semantic Componential Analysis (SCA) framework for a cross-cultural and cross-domain analysi…

Hate Speech Detection

Hate Speech Criteria: A Modular Approach to Task-Specific Hate Speech Definitions

2022-06-30 · NAACL (WOAH) 2022 7 · Urja Khurana, Ivar Vermeulen, Eric Nalisnick, Marloes van Noorloos 외

\textbf{Offensive Content Warning}: This paper contains offensive language only for providing examples that clarify this research and do not reflect the authors' opinions. Please be aware that these examples are offensiv…

Towards Legally Enforceable Hate Speech Detection for Public Forums

2023-05-23 · Chu Fei Luo, Rohan Bhambhoria, Xiaodan Zhu, Samuel Dahan

Hate speech causes widespread and deep-seated societal issues. Proper enforcement of hate speech laws is key for protecting groups of people against harmful and discriminatory language. However, determining what constitu…

Hate Speech Detection

Hate Speech Detection Using Cross-Platform Social Media Data In English and German Language

2024-10-02 · Gautam Kishore Shahi, Tim A. Majchrzak

Hate speech has grown into a pervasive phenomenon, intensifying during times of crisis, elections, and social unrest. Multiple approaches have been developed to detect hate speech using artificial intelligence, but a gen…

ClassificationHate Speech Detectiontext-classificationText Classification