paper-with-me

Papers

Multilingual HateCheck: Functional Tests for Multilingual Hate Speech Detection Models

2022-06-20 · NAACL (WOAH) 2022 7 · Paul Röttger, Haitham Seelawi, Debora Nozza, Zeerak Talat, Bertie Vidgen

Hate speech detection models are typically evaluated on held-out test sets. However, this risks painting an incomplete and potentially misleading picture of model performance because of increasingly well-documented systematic gaps and biases in hate speech datasets. To enable more targeted diagnostic insights, recent research has thus introduced functional tests for hate speech detection models. However, these tests currently only exist for English-language content, which means that they cannot support the development of more effective models in other languages spoken by billions across the world. To help address this issue, we introduce Multilingual HateCheck (MHC), a suite of functional tests for multilingual hate speech detection models. MHC covers 34 functionalities across ten languages, which is more languages than any other hate speech dataset. To illustrate MHC's utility, we train and test a high-performing multilingual hate speech detection model, and reveal critical model weaknesses for monolingual and cross-lingual applications.

📄 PDF Abstract BibTeX arXiv:2206.09917

Code (1)

rewire-online/multilingual-hatecheck 공식 구현

Tasks

DiagnosticHate Speech Detection

Methods 이 논문이 사용한 방법론

Test 설명 없음

Similar Papers 제목 키워드 기반

SEAHateCheck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of Southeast Asia

2026-03-17 · Ri Chi Ng, Aditi Kumaresan, Yujia Hu, Roy Ka-Wei Lee arxiv

Hate speech detection relies heavily on linguistic resources, which are primarily available in high-resource languages such as English and Chinese, creating barriers for researchers and platforms developing tools for low…

Hate Speech Detection

HateCheckHIn: Evaluating Hindi Hate Speech Detection Models

2022-04-30 · LREC 2022 6 · Mithun Das, Punyajoy Saha, Binny Mathew, Animesh Mukherjee

Due to the sheer volume of online hate, the AI and NLP communities have started building models to detect such hateful content. Recently, multilingual hate is a major emerging challenge for automated detection where code…

DiagnosticHate Speech Detection

GPT-HateCheck: Can LLMs Write Better Functional Tests for Hate Speech Detection?

2024-02-23 · Yiping Jin, Leo Wanner, Alexander Shvets

Online hate detection suffers from biases incurred in data sampling, annotation, and model pre-training. Therefore, measuring the averaged performance over all examples in held-out test data is inadequate. Instead, we mu…

DiagnosticHate Speech DetectionNatural Language InferenceSentence

SGHateCheck: Functional Tests for Detecting Hate Speech in Low-Resource Languages of Singapore

2024-05-03 · Ri Chi Ng, Nirmalendu Prakash, Ming Shan Hee, Kenny Tsu Wei Choo 외

To address the limitations of current hate speech detection models, we introduce \textsf{SGHateCheck}, a novel framework designed for the linguistic and cultural context of Singapore and Southeast Asia. It extends the fu…

Hate Speech DetectionTranslation

HateCheck: Functional Tests for Hate Speech Detection Models

2020-12-31 · ACL 2021 5 · Paul Röttger, Bertram Vidgen, Dong Nguyen, Zeerak Waseem 외

Detecting online hate is a difficult task that even state-of-the-art models struggle with. Typically, hate speech detection models are evaluated by measuring their performance on held-out test data using metrics such as …

DiagnosticHate Speech Detection