paper-with-me

홈 › Papers

Beyond Plain Toxic: Detection of Inappropriate Statements on Flammable Topics for the Russian Language

2022-03-04 · Nikolay Babakov, Varvara Logacheva, Alexander Panchenko

Toxicity on the Internet, such as hate speech, offenses towards particular users or groups of people, or the use of obscene words, is an acknowledged problem. However, there also exist other types of inappropriate messages which are usually not viewed as toxic, e.g. as they do not contain explicit offences. Such messages can contain covered toxicity or generalizations, incite harmful actions (crime, suicide, drug use), provoke "heated" discussions. Such messages are often related to particular sensitive topics, e.g. on politics, sexual minorities, social injustice which more often than other topics, e.g. cars or computing, yield toxic emotional reactions. At the same time, clearly not all messages within such flammable topics are inappropriate. Towards this end, in this work, we present two text collections labelled according to binary notion of inapropriateness and a multinomial notion of sensitive topic. Assuming that the notion of inappropriateness is common among people of the same culture, we base our approach on human intuitive understanding of what is not acceptable and harmful. To objectivise the notion of inappropriateness, we define it in a data-driven way though crowdsourcing. Namely we run a large-scale annotation study asking workers if a given chatbot textual statement could harm reputation of a company created it. Acceptably high values of inter-annotator agreement suggest that the notion of inappropriateness exists and can be uniformly understood by different people. To define the notion of sensitive topics in an objective way we use on guidelines suggested commonly by specialists of legal and PR department of a large public company as potentially harmful.

📄 PDF Abstract BibTeX arXiv:2203.02392

Code (0)

등록된 구현이 없습니다.

Tasks

ChatbotCultural Vocal Bursts Intensity Prediction

Methods 이 논문이 사용한 방법론

BASE 설명 없음

Similar Papers 제목 키워드 기반

Detecting Inappropriate Messages on Sensitive Topics that Could Harm a Company's Reputation

2021-03-09 · Nikolay Babakov, Varvara Logacheva, Olga Kozlova, Nikita Semenov 외

Not all topics are equally "flammable" in terms of toxicity: a calm discussion of turtles or fishing less often fuels inappropriate toxic dialogues than a discussion of politics or sexual minorities. We define a set of s…

Detecting Inappropriate Messages on Sensitive Topics that Could Harm a Company’s Reputation

2021-04-01 · EACL (BSNLP) 2021 4 · Nikolay Babakov, Varvara Logacheva, Olga Kozlova, Nikita Semenov 외

Not all topics are equally “flammable” in terms of toxicity: a calm discussion of turtles or fishing less often fuels inappropriate toxic dialogues than a discussion of politics or sexual minorities. We define a set of s…

COBRA Frames: Contextual Reasoning about Effects and Harms of Offensive Statements

2023-06-03 · Xuhui Zhou, Hao Zhu, Akhila Yerukola, Thomas Davidson 외

Warning: This paper contains content that may be offensive or upsetting. Understanding the harms and offensiveness of statements requires reasoning about the social and situational context in which statements are made. F…

Efficient Toxic Content Detection by Bootstrapping and Distilling Large Language Models

2023-12-13 · Jiang Zhang, Qiong Wu, Yiming Xu, Cheng Cao 외

Toxic content detection is crucial for online services to remove inappropriate content that violates community standards. To automate the detection process, prior works have proposed varieties of machine learning (ML) ap…

In-Context Learning

ToxiGen: A Large-Scale Machine-Generated Dataset for Adversarial and Implicit Hate Speech Detection

2022-03-17 · ACL 2022 5 · Thomas Hartvigsen, Saadia Gabriel, Hamid Palangi, Maarten Sap 외

Toxic language detection systems often falsely flag text that contains minority group mentions as toxic, as those groups are often the targets of online hate. Such over-reliance on spurious correlations also causes syste…

Hate Speech DetectionLanguage Modelling