paper-with-me

홈 › Papers

Learning to Decipher Hate Symbols

2019-04-04 · NAACL 2019 6 · Jing Qian, Mai ElSherief, Elizabeth Belding, William Yang Wang

Existing computational models to understand hate speech typically frame the problem as a simple classification task, bypassing the understanding of hate symbols (e.g., 14 words, kigy) and their secret connotations. In this paper, we propose a novel task of deciphering hate symbols. To do this, we leverage the Urban Dictionary and collected a new, symbol-rich Twitter corpus of hate speech. We investigate neural network latent context models for deciphering hate symbols. More specifically, we study Sequence-to-Sequence models and show how they are able to crack the ciphers based on context. Furthermore, we propose a novel Variational Decipher and show how it can generalize better to unseen hate symbols in a more challenging testing setting.

📄 PDF Abstract BibTeX arXiv:1904.02418

Code (0)

등록된 구현이 없습니다.

Tasks

General Classification

Similar Papers 제목 키워드 기반

Hateful Symbols or Hateful People? Predictive Features for Hate Speech Detection on Twitter

2016-06-01 · NAACL 2016 6 · Zeerak Waseem, Dirk Hovy
Hate Speech Detection

Learning to Decipher from Pixels: A Case Study of Copiale

2026-04-26 · Lei Kang, Giuseppe De Gregorio, Raphaela Heil, Alicia Fornés 외 arxiv

Historical encrypted manuscripts require both paleographic interpretation of cipher symbols and cryptanalytic recovery of plaintext. Most existing computational workflows rely on a transcription-first paradigm, in which …

Deciphering Implicit Hate: Evaluating Automated Detection Algorithms for Multimodal Hate

2021-06-10 · Findings (ACL) 2021 8 · Austin Botelho, Bertie Vidgen, Scott A. Hale

Accurate detection and classification of online hate is a difficult task. Implicit hate is particularly challenging as such content tends to have unusual syntax, polysemic words, and fewer markers of prejudice (e.g., slu…

More Than Sum of Its Parts: Deciphering Intent Shifts in Multimodal Hate Speech Detection

2026-03-22 · Runze Sun, Yu Zheng, Zexuan Xiong, Zhongjin Qu 외 arxiv

Combating hate speech on social media is critical for securing cyberspace, yet relies heavily on the efficacy of automated detection systems. As content formats evolve, hate speech is transitioning from solely plain text…

Hate Speech DetectionBinary Classification

MineriaUNAM at SemEval-2019 Task 5: Detecting Hate Speech in Twitter using Multiple Features in a Combinatorial Framework

2019-06-01 · SEMEVAL 2019 6 · Luis Enrique Argota Vega, Jorge Carlos Reyes-Maga{\~n}a, Helena G{\'o}mez-Adorno, Gemma Bel-Enguix

This paper presents our approach to the Task 5 of Semeval-2019, which aims at detecting hate speech against immigrants and women in Twitter. The task consists of two sub-tasks, in Spanish and English: (A) detection of ha…

General ClassificationPOS