paper-with-me

홈 › Papers

Studying Generalisability across Abusive Language Detection Datasets

2019-11-01 · CONLL 2019 11 · Steve Durairaj Swamy, Anupam Jamatia, Bj{\"o}rn Gamb{\"a}ck

Work on Abusive Language Detection has tackled a wide range of subtasks and domains. As a result of this, there exists a great deal of redundancy and non-generalisability between datasets. Through experiments on cross-dataset training and testing, the paper reveals that the preconceived notion of including more non-abusive samples in a dataset (to emulate reality) may have a detrimental effect on the generalisability of a model trained on that data. Hence a hierarchical annotation model is utilised here to reveal redundancies in existing datasets and to help reduce redundancy in future efforts.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Abusive Language

Similar Papers 제목 키워드 기반

Generalisability of Topic Models in Cross-corpora Abusive Language Detection

2021-06-01 · NAACL (NLP4IF) 2021 6 · Tulika Bose, Irina Illina, Dominique Fohr

Rapidly changing social media content calls for robust and generalisable abuse detection models. However, the state-of-the-art supervised models display degraded performance when they are evaluated on abusive comments th…

Abuse DetectionAbusive LanguageTopic Models

XHate-999: Analyzing and Detecting Abusive Language Across Domains and Languages

2020-12-01 · COLING 2020 8 · Goran Glava{\v{s}}, Mladen Karan, Ivan Vuli{\'c}

We present XHate-999, a multi-domain and multilingual evaluation data set for abusive language detection. By aligning test instances across six typologically diverse languages, XHate-999 for the first time allows for dis…

Abusive LanguageDisentanglementLanguage ModelingLanguage Modelling+1

Examining Temporal Bias in Abusive Language Detection

2023-09-25 · Mali Jin, Yida Mu, Diana Maynard, Kalina Bontcheva

The use of abusive language online has become an increasingly pervasive problem that damages both individuals and society, with effects ranging from psychological harm right through to escalation to real-life violence an…

Abusive Language

Detect All Abuse! Toward Universal Abusive Language Detection Models

2020-10-08 · COLING 2020 8 · Kunze Wang, Dong Lu, Soyeon Caren Han, Siqu Long 외

Online abusive language detection (ALD) has become a societal issue of increasing importance in recent years. Several previous works in online ALD focused on solving a single abusive language problem in a single domain, …

Abusive LanguageAllGraph Embedding

Data Bootstrapping Approaches to Improve Low Resource Abusive Language Detection for Indic Languages

2022-04-26 · Mithun Das, Somnath Banerjee, Animesh Mukherjee

Abusive language is a growing concern in many social media platforms. Repeated exposure to abusive speech has created physiological effects on the target users. Thus, the problem of abusive language should be addressed i…

Abusive Language