paper-with-me

홈 › Papers

Do You Really Want to Hurt Me? Predicting Abusive Swearing in Social Media

2020-05-01 · LREC 2020 5 · Endang Wahyu Pamungkas, Valerio Basile, Viviana Patti

Swearing plays an ubiquitous role in everyday conversations among humans, both in oral and textual communication, and occurs frequently in social media texts, typically featured by informal language and spontaneous writing. Such occurrences can be linked to an abusive context, when they contribute to the expression of hatred and to the abusive effect, causing harm and offense. However, swearing is multifaceted and is often used in casual contexts, also with positive social functions. In this study, we explore the phenomenon of swearing in Twitter conversations, taking the possibility of predicting the abusiveness of a swear word in a tweet context as the main investigation perspective. We developed the Twitter English corpus SWAD (Swear Words Abusiveness Dataset), where abusive swearing is manually annotated at the word level. Our collection consists of 1,511 unique swear words from 1,320 tweets. We developed models to automatically predict abusive swearing, to provide an intrinsic evaluation of SWAD and confirm the robustness of the resource. We also present the results of a glass box ablation study in order to investigate which lexical, syntactic, and affective features are more informative towards the automatic prediction of the function of swearing.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Contextual-Lexicon Approach for Abusive Language Detection

2021-04-25 · RANLP 2021 9 · Francielle Vargas, Fabiana Rodrigues de Góes, Isabelle Carvalho, Fabrício Benevenuto 외

Since a lexicon-based approach is more elegant scientifically, explaining the solution components and being easier to generalize to other applications, this paper provides a new approach for offensive language and hate s…

Abusive LanguageHate Speech Detection

Multi-word Expressions for Abusive Speech Detection in Serbian

2020-12-01 · COLING (MWE) 2020 12 · Ranka Stanković, Jelena Mitrović, Danka Jokić, Cvetana Krstev

This paper presents our work on the refinement and improvement of the Serbian language part of Hurtlex, a multilingual lexicon of words to hurt. We pay special attention to adding Multi-word expressions that can be seen …

Abusive Language

Cross-domain and Cross-lingual Abusive Language Detection: A Hybrid Approach with Deep Learning and a Multilingual Lexicon

2019-07-01 · ACL 2019 7 · Endang Wahyu Pamungkas, Viviana Patti

The development of computational methods to detect abusive language in social media within variable and multilingual contexts has recently gained significant traction. The growing interest is confirmed by the large numbe…

Abusive Language

HurtBERT: Incorporating Lexical Features with BERT for the Detection of Abusive Language

2020-11-01 · EMNLP (ALW) 2020 11 · Anna Koufakou, Endang Wahyu Pamungkas, Valerio Basile, Viviana Patti

The detection of abusive or offensive remarks in social texts has received significant attention in research. In several related shared tasks, BERT has been shown to be the state-of-the-art. In this paper, we propose to …

Abusive LanguageSentence

FlorUniTo@TRAC-2: Retrofitting Word Embeddings on an Abusive Lexicon for Aggressive Language Detection

2020-05-01 · LREC 2020 5 · Anna Koufakou, Valerio Basile, Viviana Patti

This paper describes our participation to the TRAC-2 Shared Tasks on Aggression Identification. Our team, FlorUniTo, investigated the applicability of using an abusive lexicon to enhance word embeddings towards improving…

Aggression IdentificationWord Embeddings