paper-with-me

홈 › Papers

A Machine Learning Approach to Comment Toxicity Classification

2019-02-27 · Navoneel Chakrabarty

Now-a-days, derogatory comments are often made by one another, not only in offline environment but also immensely in online environments like social networking websites and online communities. So, an Identification combined with Prevention System in all social networking websites and applications, including all the communities, existing in the digital world is a necessity. In such a system, the Identification Block should identify any negative online behaviour and should signal the Prevention Block to take action accordingly. This study aims to analyse any piece of text and detecting different types of toxicity like obscenity, threats, insults and identity-based hatred. The labelled Wikipedia Comment Dataset prepared by Jigsaw is used for the purpose. A 6-headed Machine Learning tf-idf Model has been made and trained separately, yielding a Mean Validation Accuracy of 98.08% and Absolute Validation Accuracy of 91.61%. Such an Automated System should be deployed for enhancing healthy online conversation

📄 PDF Abstract BibTeX arXiv:1903.06765

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine LearningClassificationGeneral Classification

Methods 이 논문이 사용한 방법론

Jigsaw Jigsaw is a self-supervision approach that relies on jigsaw-like puzzles as the pretext task in order to learn image representations.

Similar Papers 제목 키워드 기반

IRCologne at GermEval 2021: Toxicity Classification

2021-09-01 · GermEval 2021 9 · Fabian Haak, Björn Engelmann

In this paper, we describe the TH Köln’s submission for the Shared Task on the Identification of Toxic Comments at GermEval 2021. Toxicity is a severe and latent problem in comments in online discussions. Complex languag…

ClassificationLanguage ModelingLanguage ModellingToxic Comment Classification

AI-UPV at IberLEF-2021 DETOXIS task: Toxicity Detection in Immigration-Related Web News Comments Using Transformers and Statistical Models

2021-11-08 · Angel Felipe Magnossão de Paula, Ipek Baris Schlicht

This paper describes our participation in the DEtection of TOXicity in comments In Spanish (DETOXIS) shared task 2021 at the 3rd Workshop on Iberian Languages Evaluation Forum. The shared task is divided into two related…

ArticlesTask 2

Predicting Different Types of Subtle Toxicity in Unhealthy Online Conversations

2021-06-07 · Shlok Gilda, Mirela Silva, Luiz Giovanini, Daniela Oliveira

This paper investigates the use of machine learning models for the classification of unhealthy online conversations containing one or more forms of subtler abuse, such as hostility, sarcasm, and generalization. We levera…

Sentiment Analysis

Investigating Bias In Automatic Toxic Comment Detection: An Empirical Study

2021-08-14 · Ayush Kumar, Pratik Kumar

With surge in online platforms, there has been an upsurge in the user engagement on these platforms via comments and reactions. A large portion of such textual comments are abusive, rude and offensive to the audience. Wi…

Is Your Toxicity My Toxicity? Exploring the Impact of Rater Identity on Toxicity Annotation

2022-05-01 · Nitesh Goyal, Ian Kivlichan, Rachel Rosen, Lucy Vasserman

Machine learning models are commonly used to detect toxicity in online conversations. These models are trained on datasets annotated by human raters. We explore how raters' self-described identities impact how they annot…