paper-with-me

홈 › Papers

Hateful Symbols or Hateful People? Predictive Features for Hate Speech Detection on Twitter

2016-06-01 · NAACL 2016 6 · Zeerak Waseem, Dirk Hovy
📄 PDF Abstract BibTeX

Code (1)

zeerakw/hatespeech 공식 구현

Tasks

Hate Speech Detection

Similar Papers 제목 키워드 기반

On the Evolution of (Hateful) Memes by Means of Multimodal Contrastive Learning

2022-12-13 · Yiting Qu, Xinlei He, Shannon Pierson, Michael Backes 외

The dissemination of hateful memes online has adverse effects on social media platforms and the real world. Detecting hateful memes is challenging, one of the reasons being the evolutionary nature of memes; new hateful m…

Contrastive Learning

A Web of Hate: Tackling Hateful Speech in Online Social Spaces

2017-09-28 · Haji Mohammad Saleem, Kelly P Dillon, Susan Benesch, Derek Ruths

Online social platforms are beset with hateful speech - content that expresses hatred for a person or group of people. Such content can frighten, intimidate, or silence platform users, and some of it can inspire other us…

Now You See the Hate: Adaptive View Retrieval for Hidden Hateful Illusions

2026-07-21 · Qianpu Chen, Derya Soydaner arxiv

Hateful optical illusions expose a serious gap in current multimodal safety systems. On original-view hateful illusions, previous work shows that six moderation classifiers achieve at most 20.9 to 24.5% accuracy and nine…

Tackling Racial Bias in Automated Online Hate Detection: Towards Fair and Accurate Classification of Hateful Online Users Using Geometric Deep Learning

2021-03-22 · Zo Ahmed, Bertie Vidgen, Scott A. Hale

Online hate is a growing concern on many social media platforms and other sites. To combat it, technology companies are increasingly identifying and sanctioning `hateful users' rather than simply moderating hateful conte…

Deep LearningFairnessFeature Engineering

Unsafe Diffusion: On the Generation of Unsafe Images and Hateful Memes From Text-To-Image Models

2023-05-23 · Yiting Qu, Xinyue Shen, Xinlei He, Michael Backes 외

State-of-the-art Text-to-Image models like Stable Diffusion and DALLE$\cdot$2 are revolutionizing how people generate visual content. At the same time, society has serious concerns about how adversaries can exploit such …