paper-with-me

Papers

Recognizing Explicit and Implicit Hate Speech Using a Weakly Supervised Two-path Bootstrapping Approach

2017-10-20 · IJCNLP 2017 11 · Lei Gao, Alexis Kuppersmith, Ruihong Huang

In the wake of a polarizing election, social media is laden with hateful content. To address various limitations of supervised hate speech classification methods including corpus bias and huge cost of annotation, we propose a weakly supervised two-path bootstrapping approach for an online hate speech detection model leveraging large-scale unlabeled data. This system significantly outperforms hate speech detection systems that are trained in a supervised manner using manually annotated data. Applying this model on a large quantity of tweets collected before, after, and on election day reveals motivations and patterns of inflammatory language.

📄 PDF Abstract BibTeX arXiv:1710.07394

Code (0)

등록된 구현이 없습니다.

Tasks

General ClassificationHate Speech Detection

Similar Papers 제목 키워드 기반

Target Span Detection for Implicit Harmful Content

2024-03-28 · Nazanin Jafari, James Allan, Sheikh Muhammad Sarwar

Identifying the targets of hate speech is a crucial step in grasping the nature of such speech and, ultimately, in improving the detection of offensive posts on online forums. Much harmful content on online platforms use…

Leveraging World Knowledge in Implicit Hate Speech Detection

2022-12-28 · Jessica Lin

While much attention has been paid to identifying explicit hate speech, implicit hateful expressions that are disguised in coded or indirect language are pervasive and remain a major challenge for existing hate speech de…

Entity LinkingHate Speech DetectionWorld Knowledge

Transfer Learning via Lexical Relatedness: A Sarcasm and Hate Speech Case Study

2025-08-22 · Angelly Cabrera, Linus Lei, Antonio Ortega arxiv

Detecting hate speech in non-direct forms, such as irony, sarcasm, and innuendos, remains a persistent challenge for social networks. Although sarcasm and hate speech are regarded as distinct expressions, our work explor…

Hate Speech DetectionTransfer Learning

HatePrototypes: Interpretable and Transferable Representations for Implicit and Explicit Hate Speech Detection

2025-11-09 · Irina Proskurina, Marc-Antoine Carpentier, Julien Velcin arxiv

Optimization of offensive content moderation models for different types of hateful messages is typically achieved through continued pre-training or fine-tuning on new hate speech benchmarks. However, existing benchmarks …

Hate Speech Detection

CoSyn: Detecting Implicit Hate Speech in Online Conversations Using a Context Synergized Hyperbolic Network

2023-03-02 · Sreyan Ghosh, Manan Suri, Purva Chiniya, Utkarsh Tyagi 외

The tremendous growth of social media users interacting in online conversations has led to significant growth in hate speech, affecting people from various demographics. Most of the prior works focus on detecting explici…