paper-with-me

Papers

Large-Scale Hate Speech Detection with Cross-Domain Transfer

2021-10-16 · ACL ARR October 2021 10 · Anonymous

Hate speech towards people with different backgrounds is a major problem observed in social media. Although there are various attempts to detect hate speech automatically via supervised learning models, the performance of such models simply rely on limited datasets on which models are trained. In this study, we construct large-scale tweet datasets for supervised hate speech detection in English and Turkish, including human-labeled 100k tweets per each. Our datasets are designed to have equal number of tweets distributed over five domains; namely religion, gender, race, politics, and sports. We analyze the performance of state-of-the-art language models on large-scale hate speech detection with a special focus on model scalability. We also examine cross-domain transfer ability of hate speech detection.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Hate Speech Detection

Similar Papers 제목 키워드 기반

Large-Scale Hate Speech Detection with Cross-Domain Transfer

2022-03-02 · LREC 2022 6 · Cagri Toraman, Furkan Şahinuç, Eyup Halit Yilmaz

The performance of hate speech detection models relies on the datasets on which the models are trained. Existing datasets are mostly prepared with a limited number of instances or hate domains that define hate topics. Th…

Hate Speech DetectionTransfer Learning

Robust Hate Speech Detection in Social Media: A Cross-Dataset Empirical Evaluation

2023-07-04 · Dimosthenis Antypas, Jose Camacho-Collados

The automatic detection of hate speech online is an active research area in NLP. Most of the studies to date are based on social media datasets that contribute to the creation of hate speech detection models trained on t…

Hate Speech Detection

From BERT to Qwen: Hate Detection across architectures

2025-07-14 · Ariadna Mon, Saúl Fenollosa, Jon Lecumberri arxiv

Online platforms struggle to curb hate speech without over-censoring legitimate discourse. Early bidirectional transformer encoders made big strides, but the arrival of ultra-large autoregressive LLMs promises deeper con…

ImpliHateVid: A Benchmark Dataset and Two-stage Contrastive Learning Framework for Implicit Hate Speech Detection in Videos

2025-08-07 · Mohammad Zia Ur Rehman, Anukriti Bhatnagar, Omkar Kabde, Shubhi Bansal 외 arxiv

The existing research has primarily focused on text and image-based hate speech detection, video-based approaches remain underexplored. In this work, we introduce a novel dataset, ImpliHateVid, specifically curated for i…

Hate Speech DetectionContrastive Learning

Towards Hate Speech Detection at Large via Deep Generative Modeling

2020-05-13 · Tomer Wullach, Amir Adler, Einat Minkov

Hate speech detection is a critical problem in social media platforms, being often accused for enabling the spread of hatred and igniting physical violence. Hate speech detection requires overwhelming resources including…

DiversityHate Speech DetectionLanguage Modelling