paper-with-me

홈 › Papers

The Enemy Among Us: Detecting Hate Speech with Threats Based 'Othering' Language Embeddings

2018-01-23 · Wafa Alorainy, Pete Burnap, Han Liu, Matthew Williams

Offensive or antagonistic language targeted at individuals and social groups based on their personal characteristics (also known as cyber hate speech or cyberhate) has been frequently posted and widely circulated viathe World Wide Web. This can be considered as a key risk factor for individual and societal tension linked toregional instability. Automated Web-based cyberhate detection is important for observing and understandingcommunity and regional societal tension - especially in online social networks where posts can be rapidlyand widely viewed and disseminated. While previous work has involved using lexicons, bags-of-words orprobabilistic language parsing approaches, they often suffer from a similar issue which is that cyberhate can besubtle and indirect - thus depending on the occurrence of individual words or phrases can lead to a significantnumber of false negatives, providing inaccurate representation of the trends in cyberhate. This problemmotivated us to challenge thinking around the representation of subtle language use, such as references toperceived threats from "the other" including immigration or job prosperity in a hateful context. We propose anovel framework that utilises language use around the concept of "othering" and intergroup threat theory toidentify these subtleties and we implement a novel classification method using embedding learning to computesemantic distances between parts of speech considered to be part of an "othering" narrative. To validate ourapproach we conduct several experiments on different types of cyberhate, namely religion, disability, race andsexual orientation, with F-measure scores for classifying hateful instances obtained through applying ourmodel of 0.93, 0.86, 0.97 and 0.98 respectively, providing a significant improvement in classifier accuracy overthe state-of-the-art

📄 PDF Abstract BibTeX arXiv:1801.07495

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Investigating Deep Learning Approaches for Hate Speech Detection in Social Media

2020-05-29 · Prashant Kapil, Asif Ekbal, Dipankar Das

The phenomenal growth on the internet has helped in empowering individual's expressions, but the misuse of freedom of expression has also led to the increase of various cyber crimes and anti-social activities. Hate speec…

Hate Speech Detection

Hate speech detection in algerian dialect using deep learning

2023-09-20 · Dihia Lanasri, Juan Olano, Sifal Klioui, Sin Liang Lee 외

With the proliferation of hate speech on social networks under different formats, such as abusive language, cyberbullying, and violence, etc., people have experienced a significant increase in violence, putting them in u…

Abusive LanguageDeep LearningHate Speech Detection

HateBench: Benchmarking Hate Speech Detectors on LLM-Generated Content and Hate Campaigns

2025-01-28 · Xinyue Shen, Yixin Wu, Yiting Qu, Michael Backes 외

Large Language Models (LLMs) have raised increasing concerns about their misuse in generating hate speech. Among all the efforts to address this issue, hate speech detectors play a crucial role. However, the effectivenes…

Adversarial AttackBenchmarkingHate Speech Detection

Leveraging Transformers for Hate Speech Detection in Conversational Code-Mixed Tweets

2021-12-18 · Zaki Mustafa Farooqi, Sreyan Ghosh, Rajiv Ratn Shah

In the current era of the internet, where social media platforms are easily accessible for everyone, people often have to deal with threats, identity attacks, hate, and bullying due to their association with a cast, cree…

Hate Speech Detection

MineriaUNAM at SemEval-2019 Task 5: Detecting Hate Speech in Twitter using Multiple Features in a Combinatorial Framework

2019-06-01 · SEMEVAL 2019 6 · Luis Enrique Argota Vega, Jorge Carlos Reyes-Maga{\~n}a, Helena G{\'o}mez-Adorno, Gemma Bel-Enguix

This paper presents our approach to the Task 5 of Semeval-2019, which aims at detecting hate speech against immigrants and women in Twitter. The task consists of two sub-tasks, in Spanish and English: (A) detection of ha…

General ClassificationPOS