paper-with-me

홈 › Papers

Detecting Inappropriate Messages on Sensitive Topics that Could Harm a Company's Reputation

2021-03-09 · Nikolay Babakov, Varvara Logacheva, Olga Kozlova, Nikita Semenov, Alexander Panchenko

Not all topics are equally "flammable" in terms of toxicity: a calm discussion of turtles or fishing less often fuels inappropriate toxic dialogues than a discussion of politics or sexual minorities. We define a set of sensitive topics that can yield inappropriate and toxic messages and describe the methodology of collecting and labeling a dataset for appropriateness. While toxicity in user-generated data is well-studied, we aim at defining a more fine-grained notion of inappropriateness. The core of inappropriateness is that it can harm the reputation of a speaker. This is different from toxicity in two respects: (i) inappropriateness is topic-related, and (ii) inappropriate message is not toxic but still unacceptable. We collect and release two datasets for Russian: a topic-labeled dataset and an appropriateness-labeled dataset. We also release pre-trained classification models trained on this data.

📄 PDF Abstract BibTeX arXiv:2103.05345

Code (2)

skoltech-nlp/inappropriate-sensitive-topics 공식 구현
s-nlp/inappropriate-sensitive-topics

Similar Papers 제목 키워드 기반

Detecting Inappropriate Messages on Sensitive Topics that Could Harm a Company’s Reputation

2021-04-01 · EACL (BSNLP) 2021 4 · Nikolay Babakov, Varvara Logacheva, Olga Kozlova, Nikita Semenov 외

Not all topics are equally “flammable” in terms of toxicity: a calm discussion of turtles or fishing less often fuels inappropriate toxic dialogues than a discussion of politics or sexual minorities. We define a set of s…

Beyond Plain Toxic: Detection of Inappropriate Statements on Flammable Topics for the Russian Language

2022-03-04 · Nikolay Babakov, Varvara Logacheva, Alexander Panchenko

Toxicity on the Internet, such as hate speech, offenses towards particular users or groups of people, or the use of obscene words, is an acknowledged problem. However, there also exist other types of inappropriate messag…

ChatbotCultural Vocal Bursts Intensity Prediction

Large Language Models for Automatic Detection of Sensitive Topics

2024-09-02 · Ruoyu Wen, Stephanie Elena Crowe, Kunal Gupta, Xinyue Li 외

Sensitive information detection is crucial in content moderation to maintain safe online communities. Assisting in this traditionally manual process could relieve human moderators from overwhelming and tedious tasks, all…

Sex, drugs, and violence

2016-08-11 · Stefania Raimondo, Frank Rudzicz

Automatically detecting inappropriate content can be a difficult NLP task, requiring understanding context and innuendo, not just identifying specific keywords. Due to the large quantity of online user-generated content,…

regression

Toward Safe Evolution of Artificial Intelligence (AI) based Conversational Agents to Support Adolescent Mental and Sexual Health Knowledge Discovery

2024-04-03 · Jinkyung Park, Vivek Singh, Pamela Wisniewski

Following the recent release of various Artificial Intelligence (AI) based Conversation Agents (CAs), adolescents are increasingly using CAs for interactive knowledge discovery on sensitive topics, including mental and s…