paper-with-me

Papers

Will I Get Hate Speech Predicting the Volume of Abusive Replies before Posting in Social Media

2025-03-04 · Raneem Alharthia, Rajwa Alharthib, Ravi Shekharc, Aiqi Jiangd, Arkaitz Zubiagaa

Despite the growing body of research tackling offensive language in social media, this research is predominantly reactive, determining if content already posted in social media is abusive. There is a gap in predictive approaches, which we address in our study by enabling to predict the volume of abusive replies a tweet will receive after being posted. We formulate the problem from the perspective of a social media user asking: ``if I post a certain message on social media, is it possible to predict the volume of abusive replies it might receive?'' We look at four types of features, namely text, text metadata, tweet metadata, and account features, which also help us understand the extent to which the user or the content helps predict the number of abusive replies. This, in turn, helps us develop a model to support social media users in finding the best way to post content. One of our objectives is also to determine the extent to which the volume of abusive replies that a tweet will get are motivated by the content of the tweet or by the identity of the user posting it. Our study finds that one can build a model that performs competitively by developing a comprehensive set of features derived from the content of the message that is going to be posted. In addition, our study suggests that features derived from the user's identity do not impact model performance, hence suggesting that it is especially the content of a post that triggers abusive replies rather than who the user is.

📄 PDF Abstract BibTeX arXiv:2503.03005

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

SET Dynamic Sparse Training method where weight mask is updated randomly periodically

Similar Papers 제목 키워드 기반

Generalizing Hate Speech Detection Using Multi-Task Learning: A Case Study of Political Public Figures

2022-08-22 · Lanqin Yuan, Marian-Andrei Rizoiu

Automatic identification of hateful and abusive content is vital in combating the spread of harmful online content and its damaging effects. Most existing works evaluate models by examining the generalization error on tr…

Hate Speech DetectionMulti-Task Learning

L-HSAB: A Levantine Twitter Dataset for Hate Speech and Abusive Language

2019-08-01 · WS 2019 8 · Hala Mulki, Hatem Haddad, Chedi Bechikh Ali, Halima Alshabani

Hate speech and abusive language have become a common phenomenon on Arabic social media. Automatic hate speech and abusive detection systems can facilitate the prohibition of toxic textual contents. The complexity, infor…

Abusive LanguageHate Speech Detection

Racial Bias in Hate Speech and Abusive Language Detection Datasets

2019-05-29 · WS 2019 8 · Thomas Davidson, Debasmita Bhattacharya, Ingmar Weber

Technologies for abusive language detection are being developed and applied with little consideration of their potential biases. We examine racial bias in five different sets of Twitter data annotated for hate speech and…

Abuse DetectionAbusive Language

Intersectional Bias in Hate Speech and Abusive Language Datasets

2020-05-12 · Jae Yeon Kim, Carlos Ortiz, Sarah Nam, Sarah Santiago 외

Algorithms are widely applied to detect hate speech and abusive language in social media. We investigated whether the human-annotated data used to train these algorithms are biased. We utilized a publicly available annot…

Abuse DetectionAbusive Language

Multi-label Hate Speech and Abusive Language Detection in Indonesian Twitter

2019-08-01 · WS 2019 8 · Muhammad Okky Ibrohim, Indra Budi

Hate speech and abusive language spreading on social media need to be detected automatically to avoid conflict between citizen. Moreover, hate speech has a target, category, and level that also needs to be detected to he…

Abuse DetectionAbusive LanguageHate Speech DetectionMulti Label Text Classification+3