paper-with-me

Papers

Separating Hate Speech and Offensive Language Classes via Adversarial Debiasing

2022-07-01 · NAACL (WOAH) 2022 7 · Shuzhou Yuan, Antonis Maronikolakis, Hinrich Schütze

Research to tackle hate speech plaguing online media has made strides in providing solutions, analyzing bias and curating data. A challenging problem is ambiguity between hate speech and offensive language, causing low performance both overall and specifically for the hate speech class. It can be argued that misclassifying actual hate speech content as merely offensive can lead to further harm against targeted groups. In our work, we mitigate this potentially harmful phenomenon by proposing an adversarial debiasing method to separate the two classes. We show that our method works for English, Arabic German and Hindi, plus in a multilingual setting, improving performance over baselines.

📄 PDF Abstract BibTeX

Code (1)

shuzhouyuan/hate_speech_adversarial_debiasing 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Afaan Oromo Hate Speech Detection and Classification on Social Media

2022-06-01 · LREC 2022 6 · Teshome Mulugeta Ababu, Michael Melese Woldeyohannis

Hate and offensive speech on social media is targeted to attack an individual or group of community based on protected characteristics such as gender, ethnicity, and religion. Hate and offensive speech on social media is…

ClassificationHate Speech Detection

Overview of the HASOC Subtrack at FIRE 2021: Hate Speech and Offensive Content Identification in English and Indo-Aryan Languages

2021-12-17 · Thomas Mandl, Sandip Modha, Gautam Kishore Shahi, Hiren Madhu 외

The widespread of offensive content online such as hate speech poses a growing societal problem. AI tools are necessary for supporting the moderation process at online platforms. For the evaluation of these identificatio…

Binary ClassificationClassification

Using Transfer-based Language Models to Detect Hateful and Offensive Language Online

2020-11-01 · EMNLP (ALW) 2020 11 · Vebjørn Isaksen, Björn Gambäck

Distinguishing hate speech from non-hate offensive language is challenging, as hate speech not always includes offensive slurs and offensive language not always express hate. Here, four deep learners based on the Bidirec…

HateBR: Large expert annotated corpus of Brazilian Instagram comments for abusive language detection

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Due to the severity of the social media abusive comments in Brazil, and the lack of research in Portuguese, this paper provides the first large-scale annotated corpus of Brazilian Instagram comments for hate speech and o…

Abusive LanguageBinary Classification

QutNocturnal@HASOC'19: CNN for Hate Speech and Offensive Content Identification in Hindi Language

2020-08-28 · Md Abul Bashar, Richi Nayak

We describe our top-team solution to Task 1 for Hindi in the HASOC contest organised by FIRE 2019. The task is to identify hate speech and offensive language in Hindi. More specifically, it is a binary classification pro…

Binary Classification