paper-with-me

Papers

Machine Learning Suites for Online Toxicity Detection

2018-10-03 · David Noever

To identify and classify toxic online commentary, the modern tools of data science transform raw text into key features from which either thresholding or learning algorithms can make predictions for monitoring offensive conversations. We systematically evaluate 62 classifiers representing 19 major algorithmic families against features extracted from the Jigsaw dataset of Wikipedia comments. We compare the classifiers based on statistically significant differences in accuracy and relative execution time. Among these classifiers for identifying toxic comments, tree-based algorithms provide the most transparently explainable rules and rank-order the predictive contribution of each feature. Among 28 features of syntax, sentiment, emotion and outlier word dictionaries, a simple bad word list proves most predictive of offensive commentary.

📄 PDF Abstract BibTeX arXiv:1810.01869

Code (0)

등록된 구현이 없습니다.

Tasks

BIG-bench Machine Learning

Methods 이 논문이 사용한 방법론

Jigsaw Jigsaw is a self-supervision approach that relies on jigsaw-like puzzles as the pretext task in order to learn image representations.

Similar Papers 제목 키워드 기반

Predictively Combatting Toxicity in Health-related Online Discussions through Machine Learning

2025-05-19 · Jorge Paz-Ruza, Amparo Alonso-Betanzos, Bertha Guijarro-Berdiñas, Carlos Eiras-Franco

In health-related topics, user toxicity in online discussions frequently becomes a source of social conflict or promotion of dangerous, unscientific behaviour; common approaches for battling it include different forms of…

Collaborative Filtering

Empirical Analysis of Multi-Task Learning for Reducing Model Bias in Toxic Comment Detection

2019-09-21 · Ameya Vaidya, Feng Mai, Yue Ning

With the recent rise of toxicity in online conversations on social media platforms, using modern machine learning algorithms for toxic comment detection has become a central focus of many online applications. Researchers…

Multi-Task Learning

Using Sentiment Information for Preemptive Detection of Toxic Comments in Online Conversations

2020-06-17 · Éloi Brassard-Gourdeau, Richard Khoury

The challenge of automatic detection of toxic comments online has been the subject of a lot of research recently, but the focus has been mostly on detecting it in individual messages after they have been posted. Some aut…

Understanding Toxicity Triggers on Reddit in the Context of Singapore

2022-04-19 · Yun Yu Chong, Haewoon Kwak

While the contagious nature of online toxicity sparked increasing interest in its early detection and prevention, most of the literature focuses on the Western world. In this work, we demonstrate that 1) it is possible t…

Which one is more toxic? Findings from Jigsaw Rate Severity of Toxic Comments

2022-06-27 · TRAC (COLING) 2022 10 · Millon Madhur Das, Punyajoy Saha, Mithun Das

The proliferation of online hate speech has necessitated the creation of algorithms which can detect toxicity. Most of the past research focuses on this detection as a classification task, but assigning an absolute toxic…

regression