paper-with-me

Papers

AbuseAnalyzer: Abuse Detection, Severity and Target Prediction for Gab Posts

2020-09-30 · COLING 2020 8 · Mohit Chandra, Ashwin Pathak, Eesha Dutta, Paryul Jain, Manish Gupta, Manish Shrivastava, Ponnurangam Kumaraguru

While extensive popularity of online social media platforms has made information dissemination faster, it has also resulted in widespread online abuse of different types like hate speech, offensive language, sexist and racist opinions, etc. Detection and curtailment of such abusive content is critical for avoiding its psychological impact on victim communities, and thereby preventing hate crimes. Previous works have focused on classifying user posts into various forms of abusive behavior. But there has hardly been any focus on estimating the severity of abuse and the target. In this paper, we present a first of the kind dataset with 7601 posts from Gab which looks at online abuse from the perspective of presence of abuse, severity and target of abusive behavior. We also propose a system to address these tasks, obtaining an accuracy of ~80% for abuse presence, ~82% for abuse target prediction, and ~65% for abuse severity prediction.

📄 PDF Abstract BibTeX arXiv:2010.00038

Code (1)

mohit3011/AbuseAnalyzer 공식 구현 tf

Tasks

Abuse Detectionseverity prediction

Similar Papers 제목 키워드 기반

Deep Prompt Multi-task Network for Abuse Language Detection

2024-03-08 · Jian Zhu, YuPing Ruan, Jingfei Chang, Wenhui Sun 외

The detection of abusive language remains a long-standing challenge with the extensive use of social networks. The detection task of abusive language suffers from limited accuracy. We argue that the existing detection me…

Abusive LanguageGeneral KnowledgeMulti-Task Learning

Conditional Reliability of Toxicity Signals for Multilingual and Code-Mixed Abuse Detection

2026-07-17 · Indraveni Chebolu, Rohan Singh, Arnab Mallick, Harmesh Rana arxiv

Moderation systems increasingly rely on external toxicity tools, but those tools are unreliable under code-mixing, transliteration, slang, and language mismatch. We study the \emph{conditional reliability} of toxicity pr…

A Keyword Based Approach to Understanding the Overpenalization of Marginalized Groups by English Marginal Abuse Models on Twitter

2022-10-07 · Kyra Yee, Alice Schoenauer Sebag, Olivia Redfield, Emily Sheng 외

Harmful content detection models tend to have higher false positive rates for content from marginalized groups. In the context of marginal abuse modeling on Twitter, such disproportionate penalization poses the risk of r…

Bias DetectionFairness

A Crowd-based Evaluation of Abuse Response Strategies in Conversational Agents

2019-09-10 · WS 2019 9 · Amanda Cercas Curry, Verena Rieser

How should conversational agents respond to verbal abuse through the user? To answer this question, we conduct a large-scale crowd-sourced evaluation of abuse response strategies employed by current state-of-the-art syst…

Explainable Abuse Detection as Intent Classification and Slot Filling

2022-10-06 · Agostina Calabrese, Björn Ross, Mirella Lapata

To proactively offer social media users a safe online experience, there is a need for systems that can detect harmful posts and promptly alert platform moderators. In order to guarantee the enforcement of a consistent po…

Abuse DetectionGeneral Classificationintent-classificationIntent Classification+1