paper-with-me

홈 › Papers

Detecting Abusive Albanian

2021-07-28 · Erida Nurce, Jorgel Keci, Leon Derczynski

The ever growing usage of social media in the recent years has had a direct impact on the increased presence of hate speech and offensive speech in online platforms. Research on effective detection of such content has mainly focused on English and a few other widespread languages, while the leftover majority fail to have the same work put into them and thus cannot benefit from the steady advancements made in the field. In this paper we present \textsc{Shaj}, an annotated Albanian dataset for hate speech and offensive speech that has been constructed from user-generated content on various social media platforms. Its annotation follows the hierarchical schema introduced in OffensEval. The dataset is tested using three different classification models, the best of which achieves an F1 score of 0.77 for the identification of offensive language, 0.64 F1 score for the automatic categorization of offensive types and lastly, 0.52 F1 score for the offensive language target identification.

📄 PDF Abstract BibTeX arXiv:2107.13592

Code (0)

등록된 구현이 없습니다.

Tasks

Hate Speech Detection

Similar Papers 제목 키워드 기반

Detecting context abusiveness using hierarchical deep learning

2019-11-01 · WS 2019 11 · Ju-Hyoung Lee, Jun-U Park, Jeong-Won Cha, Yo-Sub Han

Abusive text is a serious problem in social media and causes many issues among users as the number of users and the content volume increase. There are several attempts for detecting or preventing abusive text effectively…

Deep Learning

Albanian Language Identification in Text Documents

2019-01-14 · Klesti Hoxha, Artur Baxhaku

In this work we investigate the accuracy of standard and state-of-the-art language identification methods in identifying Albanian in written text documents. A dataset consisting of news articles written in Albanian has b…

ArticlesGeneral ClassificationLanguage Identification

Exploring the Use of Lexicons to aid Deep Learning towards the Detection of Abusive Language

2019-08-01 · WS 2019 8 · Anna Koufakou, Jason Scott

Detecting abusive language is a significant research topic, which has received a lot of attention recently. Our work focused on detecting personal attacks in online conversations. State-of-the-art research on this task h…

Abusive LanguageWord Embeddings

Implicitly Abusive Comparisons -- A New Dataset and Linguistic Analysis

2021-04-01 · EACL 2021 2 · Michael Wiegand, Maja Geulig, Josef Ruppenhofer

We examine the task of detecting implicitly abusive comparisons (e.g. {``}Your hair looks like you have been electrocuted{''}). Implicitly abusive comparisons are abusive comparisons in which abusive words (e.g. {``}dumb…

Unraveling the Search Space of Abusive Language in Wikipedia with Dynamic Lexicon Acquisition

2019-11-01 · WS 2019 11 · Wei-Fan Chen, Khalid Al Khatib, Matthias Hagen, Henning Wachsmuth 외

Many discussions on online platforms suffer from users offending others by using abusive terminology, threatening each other, or being sarcastic. Since an automatic detection of abusive language can support human moderat…

Abusive Language