Hate Speech Detection
15개 벤치마크 · 논문 583편 · 이 태스크의 논문 보기 →
Benchmarks
Ethos Binary
HateXplain
Ethos MultiLabel
Waseem et al., 2018
HateMM
AbusEval
HatEval
OffensEval 2019
ToLD-Br
DKhate
OLID
SHAJ
bajer_danish_misogyny
Most implemented
DistilBERT, a distilled version of BERT: smaller, faster, cheaper and lighter
OPT: Open Pre-trained Transformer Language Models
Automated Hate Speech Detection and the Problem of Offensive Language
HateXplain: A Benchmark Dataset for Explainable Hate Speech Detection
Comparative Studies of Detecting Abusive Language on Twitter
Papers
From Specialization to Generalization: Instruction-tuned LLMs for Robust Harmful Content Mitigation
Large language models (LLMs) demonstrate impressive performance across a wide range of general NLP tasks; however, their effectiveness in sensitive domains, such as hate speech detection, remains less clear. Prior studie…
Hate Speech Detection'Ghaib in Translation' aka Unseen Harm: Measuring Cross-Script Safety Inconsistency with 'Missed-in-Urdu' Scores in LLM Hate Speech Detection
Urdu, the world's tenth most spoken language with 246 million speakers, remains almost entirely absent from mainstream LLM safety evaluation and nine years of WOAH proceedings. To investigate whether this absence has mea…
Hate Speech DetectionIntroducing the Privacy-HSD Trade-off: Hate Speech Detection, but not at the Cost of Privacy
Hate speech is a real and timely threat that affects a large portion of online users, especially youth and minority groups. While building reliable and robust automatic hate speech detection (HSD) systems is paramount, w…
Hate Speech DetectionTranslation as a Computationally Efficient Bridge: Feasibility of English BERT for Low-Resource Languages
BERT models have revolutionised Natural Language Processing (NLP) through their ability to process unstructured text across diverse domains. However, developing high-quality BERT models for non-English languages remains …
Natural Language InferencePart-Of-Speech TaggingHate Speech DetectionQuestion AnsweringHate Speech Detection in Turkish and Arabic: A Comprehensive Study
Online hate speech has been linked to a global rise in violence against minorities, including incidents such as mass shootings, lynchings, and ethnic cleansing. Societies grappling with this issue, particularly when hate…
Hate Speech DetectionTeamHerald@CHIPSAL 2026: Hate Speech Detection and Sentiment Analysis of Nepali Memes using Transformer-based Architectures and Ensemble Learning
The analysis of internet memes in the Nepali language is complicated by frequent code-mixing and a lack of established baseline resources. While memes inherently combine visual and textual elements, this study focuses on…
Hate Speech DetectionBinary ClassificationSentiment AnalysisEnsemble Learning