paper-with-me

홈 › Papers

Hate Speech Criteria: A Modular Approach to Task-Specific Hate Speech Definitions

2022-06-30 · NAACL (WOAH) 2022 7 · Urja Khurana, Ivar Vermeulen, Eric Nalisnick, Marloes van Noorloos, Antske Fokkens

\textbf{Offensive Content Warning}: This paper contains offensive language only for providing examples that clarify this research and do not reflect the authors' opinions. Please be aware that these examples are offensive and may cause you distress. The subjectivity of recognizing \textit{hate speech} makes it a complex task. This is also reflected by different and incomplete definitions in NLP. We present \textit{hate speech} criteria, developed with perspectives from law and social science, with the aim of helping researchers create more precise definitions and annotation guidelines on five aspects: (1) target groups, (2) dominance, (3) perpetrator characteristics, (4) type of negative group reference, and the (5) type of potential consequences/effects. Definitions can be structured so that they cover a more broad or more narrow phenomenon. As such, conscious choices can be made on specifying criteria or leaving them open. We argue that the goal and exact task developers have in mind should determine how the scope of \textit{hate speech} is defined. We provide an overview of the properties of English datasets from \url{hatespeechdata.com} that may help select the most suitable dataset for a specific scenario.

📄 PDF Abstract BibTeX arXiv:2206.15455

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

AWARE We propose to theoretically and empirically examine the effect of incorporating weighting schemes into walk-aggregating GNNs. To this end, we propose a simple, interpretable, and…

Similar Papers 제목 키워드 기반

A Modular Taxonomy for Hate Speech Definitions and Its Impact on Zero-Shot LLM Classification Performance

2025-06-23 · Matteo Melis, Gabriella Lapesa, Dennis Assenmacher

Detecting harmful content is a crucial task in the landscape of NLP applications for Social Good, with hate speech being one of its most dangerous forms. But what do we mean by hate speech, how can we define it, and how …

Specificity

Speech Detection Task Against Asian Hate: BERT the Central, While Data-Centric Studies the Crucial

2022-06-05 · Xin Lian

With the COVID-19 pandemic continuing, hatred against Asians is intensifying in countries outside Asia, especially among the Chinese. There is an urgent need to detect and prevent hate speech towards Asians effectively. …

Revisiting Hate Speech Benchmarks: From Data Curation to System Deployment

2023-06-01 · Atharva Kulkarni, Sarah Masud, Vikram Goyal, Tanmoy Chakraborty

Social media is awash with hateful content, much of which is often veiled with linguistic and topical diversity. The benchmark datasets used for hate speech detection do not account for such divagation as they are predom…

BenchmarkingHate Speech DetectionMixture-of-Experts

xList-Hate: A Checklist-Based Framework for Interpretable and Generalizable Hate Speech Detection

2026-02-05 · Adrián Girón, Pablo Miralles, Javier Huertas-Tato, Sergio D'Antonio 외 arxiv

Hate speech detection is commonly framed as a direct binary classification problem despite being a composite concept defined through multiple interacting factors that vary across legal frameworks, platform policies, and …

Hate Speech DetectionBinary Classification

Detection of Hate Speech using BERT and Hate Speech Word Embedding with Deep Model

2021-11-02 · Hind Saleh, Areej Alhothali, Kawthar Moria

The enormous amount of data being generated on the web and social media has increased the demand for detecting online hate speech. Detecting hate speech will reduce their negative impact and influence on others. A lot of…

Binary ClassificationHate Speech DetectionLanguage ModelingLanguage Modelling+1