paper-with-me

홈 › Papers

Towards Automatic Online Hate Speech Intervention Generation using Pretrained Language Model

2020-10-19 · Raj Ratn Pranesh, Ambesh Shekhar, Anish Kumar

Social media harbours substantial toxic and hateful conversations today. Curbing them has emerged as a critical challenge for governments and organizations globally. Prior research has primarily concentrated on the detection of online hate speech while ignoring further action needed to discourage individuals from using hate speech in the future. Counterspeech is an effective way to tackle online hate, leaving freedom of speech untouched. The focus is to directly intervene in the conversation with textual responses that counter the hate content and prevent it from further spreading. In this paper, we propose a novel natural language generation task for hate speech intervention, where the goal is to automatically generate responses to intervene during online conversations that contain hate speech. We sequentially analyzed the performance and capability of various state-of-the-art pretrained language models dialogue generation model for automated hate speech intervention system using automatic metric and manual human evaluation. The results indicate that the generated intervention responses are very promising in terms of relevance and contextual meaning

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Dialogue GenerationLanguage ModelingLanguage ModellingText Generation

Similar Papers 제목 키워드 기반

A Benchmark Dataset for Learning to Intervene in Online Hate Speech

2019-09-10 · IJCNLP 2019 11 · Jing Qian, Anna Bethke, Yinyin Liu, Elizabeth Belding 외

Countering online hate speech is a critical yet challenging task, but one which can be aided by the use of Natural Language Processing (NLP) techniques. Previous research has primarily focused on the development of NLP m…

Response Generation

Towards Automatic Generation of Messages Countering Online Hate Speech and Microaggressions

2022-07-01 · NAACL (WOAH) 2022 7 · Mana Ashida, Mamoru Komachi

With the widespread use of social media, online hate is increasing, and microaggressions are receiving attention. We explore the potential for using pretrained language models to automatically generate messages that comb…

Informativeness

Human-Machine Collaboration Approaches to Build a Dialogue Dataset for Hate Speech Countering

2022-11-07 · Helena Bonaldi, Sara Dellantonio, Serra Sinem Tekiroglu, Marco Guerini

Fighting online hate speech is a challenge that is usually addressed using Natural Language Processing via automatic detection and removal of hate content. Besides this approach, counter narratives have emerged as an eff…

Text Generation

Countering Online Hate Speech: An NLP Perspective

2021-09-07 · Mudit Chaudhary, Chandni Saxena, Helen Meng

Online hate speech has caught everyone's attention from the news related to the COVID-19 pandemic, US elections, and worldwide protests. Online toxicity - an umbrella term for online hateful behavior, manifests itself in…

Towards Knowledge-Grounded Counter Narrative Generation for Hate Speech

2021-06-22 · Findings (ACL) 2021 8 · Yi-Ling Chung, Serra Sinem Tekiroglu, Marco Guerini

Tackling online hatred using informed textual responses - called counter narratives - has been brought under the spotlight recently. Accordingly, a research line has emerged to automatically generate counter narratives i…