paper-with-me

Papers

KTCR: Improving Implicit Hate Detection with Knowledge Transfer driven Concept Refinement

2024-10-20 · Samarth Garg, Vivek Hruday Kavuri, Gargi Shroff, Rahul Mishra

The constant shifts in social and political contexts, driven by emerging social movements and political events, lead to new forms of hate content and previously unrecognized hate patterns that machine learning models may not have captured. Some recent literature proposes data augmentation-based techniques to enrich existing hate datasets by incorporating samples that reveal new implicit hate patterns. This approach aims to improve the model's performance on out-of-domain implicit hate instances. It is observed, that further addition of more samples for augmentation results in the decrease of the performance of the model. In this work, we propose a Knowledge Transfer-driven Concept Refinement method that distills and refines the concepts related to implicit hate samples through novel prototype alignment and concept losses, alongside data augmentation based on concept activation vectors. Experiments with several publicly available datasets show that incorporating additional implicit samples reflecting new hate patterns through concept refinement enhances the model's performance, surpassing baseline results while maintaining cross-dataset generalization capabilities.

📄 PDF Abstract BibTeX arXiv:2410.15314

Code (0)

등록된 구현이 없습니다.

Tasks

Data AugmentationTransfer Learning

Similar Papers 제목 키워드 기반

HatePrototypes: Interpretable and Transferable Representations for Implicit and Explicit Hate Speech Detection

2025-11-09 · Irina Proskurina, Marc-Antoine Carpentier, Julien Velcin arxiv

Optimization of offensive content moderation models for different types of hateful messages is typically achieved through continued pre-training or fine-tuning on new hate speech benchmarks. However, existing benchmarks …

Hate Speech Detection

Transfer Learning via Lexical Relatedness: A Sarcasm and Hate Speech Case Study

2025-08-22 · Angelly Cabrera, Linus Lei, Antonio Ortega arxiv

Detecting hate speech in non-direct forms, such as irony, sarcasm, and innuendos, remains a persistent challenge for social networks. Although sarcasm and hate speech are regarded as distinct expressions, our work explor…

Hate Speech DetectionTransfer Learning

Leveraging World Knowledge in Implicit Hate Speech Detection

2022-12-28 · Jessica Lin

While much attention has been paid to identifying explicit hate speech, implicit hateful expressions that are disguised in coded or indirect language are pervasive and remain a major challenge for existing hate speech de…

Entity LinkingHate Speech DetectionWorld Knowledge

ImpliHateVid: A Benchmark Dataset and Two-stage Contrastive Learning Framework for Implicit Hate Speech Detection in Videos

2025-08-07 · Mohammad Zia Ur Rehman, Anukriti Bhatnagar, Omkar Kabde, Shubhi Bansal 외 arxiv

The existing research has primarily focused on text and image-based hate speech detection, video-based approaches remain underexplored. In this work, we introduce a novel dataset, ImpliHateVid, specifically curated for i…

Hate Speech DetectionContrastive Learning

Multilingual Auxiliary Tasks Training: Bridging the Gap between Languages for Zero-Shot Transfer of Hate Speech Detection Models

2022-10-24 · Syrielle Montariol, Arij Riabi, Djamé Seddah

Zero-shot cross-lingual transfer learning has been shown to be highly challenging for tasks involving a lot of linguistic specificities or when a cultural gap is present between languages, such as in hate speech detectio…

Cross-Lingual TransferHate Speech Detectionnamed-entity-recognitionNamed Entity Recognition+4