paper-with-me

Papers

Are Chess Discussions Racist? An Adversarial Hate Speech Data Set

2020-11-20 · Rupak Sarkar, Ashiqur R. KhudaBukhsh

On June 28, 2020, while presenting a chess podcast on Grandmaster Hikaru Nakamura, Antonio Radi\'c's YouTube handle got blocked because it contained "harmful and dangerous" content. YouTube did not give further specific reason, and the channel got reinstated within 24 hours. However, Radi\'c speculated that given the current political situation, a referral to "black against white", albeit in the context of chess, earned him this temporary ban. In this paper, via a substantial corpus of 681,995 comments, on 8,818 YouTube videos hosted by five highly popular chess-focused YouTube channels, we ask the following research question: \emph{how robust are off-the-shelf hate-speech classifiers to out-of-domain adversarial examples?} We release a data set of 1,000 annotated comments where existing hate speech classifiers misclassified benign chess discussions as hate speech. We conclude with an intriguing analogy result on racial bias with our findings pointing out to the broader challenge of color polysemy.

📄 PDF Abstract BibTeX arXiv:2011.10280

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Detecting Hate Speech with GPT-3

2021-03-23 · Ke-Li Chiu, Annie Collins, Rohan Alexander

Sophisticated language models such as OpenAI's GPT-3 can generate hateful text that targets marginalized groups. Given this capacity, we are interested in whether large language models can be used to identify hate speech…

Few-Shot LearningHate Speech DetectionOne-Shot Learning

Are You a Racist or Am I Seeing Things? Annotator Influence on Hate Speech Detection on Twitter

2016-11-01 · WS 2016 11 · Zeerak Waseem
Hate Speech Detection

Predictive Embeddings for Hate Speech Detection on Twitter

2018-09-27 · WS 2018 10 · Rohan Kshirsagar, Tyus Cukuvac, Kathleen McKeown, Susan McGregor

We present a neural-network based approach to classifying online hate speech in general, as well as racist and sexist speech in particular. Using pre-trained word embeddings and max/mean pooling from simple, fully-connec…

Hate Speech DetectionWord Embeddings

Multilingual Cross-domain Perspectives on Online Hate Speech

2018-09-11 · Tom De Smedt, Sylvia Jaki, Eduan Kotzé, Leïla Saoud 외

In this report, we present a study of eight corpora of online hate speech, by demonstrating the NLP techniques that we used to collect and analyze the jihadist, extremist, racist, and sexist content. Analysis of the mult…

General Classificationtext-classificationText Classification

Automated Hate Speech Detection and the Problem of Offensive Language

2017-03-11 · Thomas Davidson, Dana Warmsley, Michael Macy, Ingmar Weber

A key challenge for automatic hate-speech detection on social media is the separation of hate speech from other instances of offensive language. Lexical detection methods tend to have low precision because they classify …

Hate Speech Detection