paper-with-me

Hate Speech Detection

15개 벤치마크 · 논문 583편 · 이 태스크의 논문 보기 →

Benchmarks

Ethos Binary

결과 24개

HateXplain

결과 22개

Ethos MultiLabel

결과 10개

Waseem et al., 2018

결과 6개

HateMM

결과 5개

AbusEval

결과 4개

HatEval

결과 4개

OffensEval 2019

결과 4개

ToLD-Br

결과 4개

DKhate

결과 2개

OLID

결과 2개

SHAJ

결과 2개

bajer_danish_misogyny

결과 2개

Most implemented

Papers

From Specialization to Generalization: Instruction-tuned LLMs for Robust Harmful Content Mitigation

2026-08-26 · Lukas Edman, Daryna Dementieva, Alexander Fraser arxiv

Large language models (LLMs) demonstrate impressive performance across a wide range of general NLP tasks; however, their effectiveness in sensitive domains, such as hate speech detection, remains less clear. Prior studie…

Hate Speech Detection

'Ghaib in Translation' aka Unseen Harm: Measuring Cross-Script Safety Inconsistency with 'Missed-in-Urdu' Scores in LLM Hate Speech Detection

2026-08-25 · Fawzia Zehra, Kara-Isitt, Sonal Khosla, Stephen Swift arxiv

Urdu, the world's tenth most spoken language with 246 million speakers, remains almost entirely absent from mainstream LLM safety evaluation and nine years of WOAH proceedings. To investigate whether this absence has mea…

Hate Speech Detection

Introducing the Privacy-HSD Trade-off: Hate Speech Detection, but not at the Cost of Privacy

2026-08-19 · Stephen Meisenbacher, Vlad Garbuz, Chirill Donos, Maxim Dnestreanschii 외 arxiv

Hate speech is a real and timely threat that affects a large portion of online users, especially youth and minority groups. While building reliable and robust automatic hate speech detection (HSD) systems is paramount, w…

Hate Speech Detection

Translation as a Computationally Efficient Bridge: Feasibility of English BERT for Low-Resource Languages

2026-07-14 · Hielke Muizelaar, Giulia Rivetti, Marco Spruit, Marcel Haas arxiv

BERT models have revolutionised Natural Language Processing (NLP) through their ability to process unstructured text across diverse domains. However, developing high-quality BERT models for non-English languages remains …

Natural Language InferencePart-Of-Speech TaggingHate Speech DetectionQuestion Answering

Hate Speech Detection in Turkish and Arabic: A Comprehensive Study

2026-06-30 · Somaiyeh Dehghan, Gökçe Uludoğan, Mehmet Umut Şen, Elif Erol 외 arxiv

Online hate speech has been linked to a global rise in violence against minorities, including incidents such as mass shootings, lynchings, and ethnic cleansing. Societies grappling with this issue, particularly when hate…

Hate Speech Detection

TeamHerald@CHIPSAL 2026: Hate Speech Detection and Sentiment Analysis of Nepali Memes using Transformer-based Architectures and Ensemble Learning

2026-06-07 · Ashish Acharya, Anish Khatiwada, Rohit Khadka, Pragya Aryal arxiv

The analysis of internet memes in the Nepali language is complicated by frequent code-mixing and a lack of established baseline resources. While memes inherently combine visual and textual elements, this study focuses on…

Hate Speech DetectionBinary ClassificationSentiment AnalysisEnsemble Learning

전체 583편 보기 →