paper-with-me

Papers

Towards Weakly-Supervised Hate Speech Classification Across Datasets

2023-05-04 · Yiping Jin, Leo Wanner, Vishakha Laxman Kadam, Alexander Shvets

As pointed out by several scholars, current research on hate speech (HS) recognition is characterized by unsystematic data creation strategies and diverging annotation schemata. Subsequently, supervised-learning models tend to generalize poorly to datasets they were not trained on, and the performance of the models trained on datasets labeled using different HS taxonomies cannot be compared. To ease this problem, we propose applying extremely weak supervision that only relies on the class name rather than on class samples from the annotated data. We demonstrate the effectiveness of a state-of-the-art weakly-supervised text classification model in various in-dataset and cross-dataset settings. Furthermore, we conduct an in-depth quantitative and qualitative analysis of the source of poor generalizability of HS classification models.

📄 PDF Abstract BibTeX arXiv:2305.02637

Code (0)

등록된 구현이 없습니다.

Tasks

Classificationtext-classificationText Classification

Similar Papers 제목 키워드 기반

Recognizing Explicit and Implicit Hate Speech Using a Weakly Supervised Two-path Bootstrapping Approach

2017-10-20 · IJCNLP 2017 11 · Lei Gao, Alexis Kuppersmith, Ruihong Huang

In the wake of a polarizing election, social media is laden with hateful content. To address various limitations of supervised hate speech classification methods including corpus bias and huge cost of annotation, we prop…

General ClassificationHate Speech Detection

Cross-Platform Hate Speech Detection with Weakly Supervised Causal Disentanglement

2024-04-17 · Paras Sheth, Tharindu Kumarage, Raha Moraffah, Aman Chadha 외

Content moderation faces a challenging task as social media's ability to spread hate speech contrasts with its role in promoting global connectivity. With rapidly evolving slang and hate speech, the adaptability of conve…

DisentanglementHate Speech Detection

MultiHateLoc: Towards Temporal Localisation of Multimodal Hate Content in Online Videos

2025-12-11 · Qiyue Sun, Tailin Chen, Yinghui Zhang, Yuchen Zhang 외 arxiv

The rapid growth of video content on platforms such as TikTok and YouTube has intensified the spread of multimodal hate speech, where harmful cues emerge subtly and asynchronously across visual, acoustic, and textual str…

xList-Hate: A Checklist-Based Framework for Interpretable and Generalizable Hate Speech Detection

2026-02-05 · Adrián Girón, Pablo Miralles, Javier Huertas-Tato, Sergio D'Antonio 외 arxiv

Hate speech detection is commonly framed as a direct binary classification problem despite being a composite concept defined through multiple interacting factors that vary across legal frameworks, platform policies, and …

Hate Speech DetectionBinary Classification

Label Propagation-Based Semi-Supervised Learning for Hate Speech Classification

2020-11-01 · EMNLP (insights) 2020 11 · Ashwin Geet D’Sa, Irina Illina, Dominique Fohr, Dietrich Klakow 외

Research on hate speech classification has received increased attention. In real-life scenarios, a small amount of labeled hate speech data is available to train a reliable classifier. Semi-supervised learning takes adva…

Classification