paper-with-me

홈 › Papers

APEACH: Attacking Pejorative Expressions with Analysis on Crowd-Generated Hate Speech Evaluation Datasets

2022-02-25 · Kichang Yang, Wonjun Jang, Won Ik Cho

In hate speech detection, developing training and evaluation datasets across various domains is the critical issue. Whereas, major approaches crawl social media texts and hire crowd-workers to annotate the data. Following this convention often restricts the scope of pejorative expressions to a single domain lacking generalization. Sometimes domain overlap between training corpus and evaluation set overestimate the prediction performance when pretraining language models on low-data language. To alleviate these problems in Korean, we propose APEACH that asks unspecified users to generate hate speech examples followed by minimal post-labeling. We find that APEACH can collect useful datasets that are less sensitive to the lexical overlaps between the pretraining corpus and the evaluation set, thereby properly measuring the model performance.

📄 PDF Abstract BibTeX arXiv:2202.12459

Code (1)

jason9693/apeach 공식 구현

Tasks

Hate Speech Detection

Methods 이 논문이 사용한 방법론

Refunds@Expedia|||How do I get a full refund from Expedia? “How do I get a full refund from Expedia? How do I get a full refund from Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Quick Help &…
Attention 설명 없음
Linear Layer A Linear Layer is a projection $\mathbf{XW + b}$.
WordPiece 설명 없음
Residual Connection 설명 없음
Layer Normalization Unlike batch normalization, Layer Normalization directly estimates the normalization statistics from the summed inputs…
Dropout Dropout is a regularization technique for neural networks that drops a unit (along with connections) at training time with a specified probability $p$ (a common value is…
Attention Dropout Attention Dropout is a type of dropout used in attention-based architectures, where elements are randomly dropped out of the…

Similar Papers 제목 키워드 기반

Illegal is not a Noun: Linguistic Form for Detection of Pejorative Nominalizations

2017-08-01 · WS 2017 8 · Alexis Palmer, Melissa Robinson, Kristy K. Phillips

This paper focuses on a particular type of abusive language, targeting expressions in which typically neutral adjectives take on pejorative meaning when used as nouns - compare {`}gay people{'} to {`}the gays{'}. We firs…

Abusive LanguageFormPOS

A Computational Exploration of Pejorative Language in Social Media

2021-11-01 · Findings (EMNLP) 2021 11 · Liviu P. Dinu, Ioan-Bogdan Iordache, Ana Sabina Uban, Marcos Zampieri

In this paper we study pejorative language, an under-explored topic in computational linguistics. Unlike existing models of offensive language and hate speech, pejorative language manifests itself primarily at the lexica…

Word Sense Disambiguation

PejorativITy: Disambiguating Pejorative Epithets to Improve Misogyny Detection in Italian Tweets

2024-04-03 · Arianna Muti, Federico Ruggeri, Cagri Toraman, Lorenzo Musetti 외

Misogyny is often expressed through figurative language. Some neutral words can assume a negative connotation when functioning as pejorative epithets. Disambiguating the meaning of such terms might help the detection of …

SentenceWord EmbeddingsWord Sense Disambiguation

Enriching Abusive Language Detection with Community Context

2022-06-16 · NAACL (WOAH) 2022 7 · Jana Kurrek, Haji Mohammad Saleem, Derek Ruths

Uses of pejorative expressions can be benign or actively empowering. When models for abuse detection misclassify these expressions as derogatory, they inadvertently censor productive conversations held by marginalized gr…

Abuse DetectionAbusive Language

Multilingual offensive lexicon annotated with contextual information

2021-11-16 · ACL ARR November 2021 11 · Anonymous

Online hate speech and offensive comments detection is not a trivial research problem since pragmatic (contextual) factors influence what is considered offensive. Moreover, offensive terms are hardly found in classical l…

Abusive Language