paper-with-me

Papers

HateMM: A Multi-Modal Dataset for Hate Video Classification

2023-05-06 · Mithun Das, Rohit Raj, Punyajoy Saha, Binny Mathew, Manish Gupta, Animesh Mukherjee

Hate speech has become one of the most significant issues in modern society, having implications in both the online and the offline world. Due to this, hate speech research has recently gained a lot of traction. However, most of the work has primarily focused on text media with relatively little work on images and even lesser on videos. Thus, early stage automated video moderation techniques are needed to handle the videos that are being uploaded to keep the platform safe and healthy. With a view to detect and remove hateful content from the video sharing platforms, our work focuses on hate video detection using multi-modalities. To this end, we curate ~43 hours of videos from BitChute and manually annotate them as hate or non-hate, along with the frame spans which could explain the labelling decision. To collect the relevant videos we harnessed search keywords from hate lexicons. We observe various cues in images and audio of hateful videos. Further, we build deep learning multi-modal models to classify the hate videos and observe that using all the modalities of the videos improves the overall hate speech detection performance (accuracy=0.798, macro F1-score=0.790) by ~5.7% compared to the best uni-modal model in terms of macro F1 score. In summary, our work takes the first step toward understanding and modeling hateful videos on video hosting platforms such as BitChute.

📄 PDF Abstract BibTeX arXiv:2305.03915

Code (1)

hate-alert/hatemm 공식 구현 pytorch

Tasks

ClassificationHate Speech DetectionVideo Classification

Similar Papers 제목 키워드 기반

CrisisHateMM: Multimodal Analysis of Directed and Undirected Hate Speech in Text-Embedded Images from Russia-Ukraine Conflict

2023-06-01 · IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) Workshop 2023 6 · Aashish Bhandari, Siddhant B. Shah, Surendrabikram Thapa, Usman Naseem 외

Text-embedded images are frequently used on social media to convey opinions and emotions, but they can also be a medium for disseminating hate speech, propaganda, and extremist ideologies. During the Russia-Ukraine war, …

Hate Speech Detection CrisisHateMM Benchmark

ImpliHateVid: A Benchmark Dataset and Two-stage Contrastive Learning Framework for Implicit Hate Speech Detection in Videos

2025-08-07 · Mohammad Zia Ur Rehman, Anukriti Bhatnagar, Omkar Kabde, Shubhi Bansal 외 arxiv

The existing research has primarily focused on text and image-based hate speech detection, video-based approaches remain underexplored. In this work, we introduce a novel dataset, ImpliHateVid, specifically curated for i…

Hate Speech DetectionContrastive Learning

Towards a Robust Framework for Multimodal Hate Detection: A Study on Video vs. Image-based Content

2025-02-11 · Girish A. Koushik, Diptesh Kanojia, Helen Treharne

Social media platforms enable the propagation of hateful content across different modalities such as textual, auditory, and visual, necessitating effective detection methods. While recent approaches have shown promise in…

Hate Speech DetectionVideo Classification

Revealing Temporal Label Noise in Multimodal Hateful Video Classification

2025-08-06 · Shuonan Yang, Tailin Chen, Rahul Singh, Jiangbei Yue 외 arxiv

The rapid proliferation of online multimedia content has intensified the spread of hate speech, presenting critical societal and regulatory challenges. While recent work has advanced multimodal hateful video detection, m…

Video Classification

Towards Training-free Multimodal Hate Localisation with Large Language Models

2026-02-10 · Yueming Sun, Long Yang, Jianbo Jiao, Zeyu Fu arxiv

The proliferation of hateful content in online videos poses severe threats to individual well-being and societal harmony. However, existing solutions for video hate detection either rely heavily on large-scale human anno…