The performance of multiple language models in identifying offensive language on social media
Text classification is an important topic in the field of natural language processing. It has been preliminarily applied in information retrieval, digital library, automatic abstracting, text filtering, word semantic discrimination and many other fields. The aim of this research is to use a variety of algorithms to test the ability to identify offensive posts and evaluate their performance against a variety of assessment methods. The motivation for this project is to reduce the harm of these languages to human censors by automating the screening of offending posts. The field is a new one, and despite much interest in the past two years, there has been no focus on the object of the offence. Through the experiment of this project, it should inspire future research on identification methods as well as identification content.
Code (0)
등록된 구현이 없습니다.
Tasks
Information RetrievalRetrievaltext-classificationText ClassificationMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
FBERT: A Neural Transformer for Identifying Offensive Content
Transformer-based models such as BERT, XLNET, and XLM-R have achieved state-of-the-art performance across various NLP tasks including the identification of offensive language and hate speech, an important problem in soci…
Language IdentificationXLM-RAutomatic Expansion and Retargeting of Arabic Offensive Language Training
Rampant use of offensive language on social media led to recent efforts on automatic identification of such language. Though offensive language has general characteristics, attacks on specific entities may exhibit distin…
ConvAI at SemEval-2019 Task 6: Offensive Language Identification and Categorization with Perspective and BERT
This paper presents the application of two strong baseline systems for toxicity detection and evaluates their performance in identifying and categorizing offensive language in social media. PERSPECTIVE is an API, that se…
Language IdentificationA Federated Learning Approach to Privacy Preserving Offensive Language Identification
The spread of various forms of offensive speech online is an important concern in social media. While platforms have been investing heavily in ways of coping with this problem, the question of privacy remains largely una…
Federated LearningLanguage IdentificationPrivacy PreservingOffense Detection in Dravidian Languages using Code-Mixing Index based Focal Loss
Over the past decade, we have seen exponential growth in online content fueled by social media platforms. Data generation of this scale comes with the caveat of insurmountable offensive content in it. The complexity of i…