paper-with-me

Papers

Hate Cannot Drive out Hate: Forecasting Conversation Incivility following Replies to Hate Speech

2023-12-08 · Xinchen Yu, Eduardo Blanco, Lingzi Hong

User-generated replies to hate speech are promising means to combat hatred, but questions about whether they can stop incivility in follow-up conversations linger. We argue that effective replies stop incivility from emerging in follow-up conversations - replies that elicit more incivility are counterproductive. This study introduces the task of predicting the incivility of conversations following replies to hate speech. We first propose a metric to measure conversation incivility based on the number of civil and uncivil comments as well as the unique authors involved in the discourse. Our metric approximates human judgments more accurately than previous metrics. We then use the metric to evaluate the outcomes of replies to hate speech. A linguistic analysis uncovers the differences in the language of replies that elicit follow-up conversations with high and low incivility. Experimental results show that forecasting incivility is challenging. We close with a qualitative analysis shedding light into the most common errors made by the best model.

📄 PDF Abstract BibTeX arXiv:2312.04804

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Echoes of Discord: Forecasting Hater Reactions to Counterspeech

2025-01-27 · Xiaoying Song, Sharon Lisseth Perez, Xinchen Yu, Eduardo Blanco 외

Hate speech (HS) erodes the inclusiveness of online users and propagates negativity and division. Counterspeech has been recognized as a way to mitigate the harmful consequences. While some research has investigated the …

Dialogues of Dissent: Thematic and Rhetorical Dimensions of Hate and Counter-Hate Speech in Social Media Conversations

2025-07-28 · Effi Levi, Gal Ron, Odelia Oshri, Shaul R. Shenhav arxiv

We introduce a novel multi-labeled scheme for joint annotation of hate and counter-hate speech in social media conversations, categorizing hate and counter-hate messages into thematic and rhetorical dimensions. The thema…

Outcome-Constrained Large Language Models for Countering Hate Speech

2024-03-25 · Lingzi Hong, Pengcheng Luo, Eduardo Blanco, Xiaoying Song

Automatic counterspeech generation methods have been developed to assist efforts in combating hate speech. Existing research focuses on generating counterspeech with linguistic attributes such as being polite, informativ…

Reinforcement Learning (RL)Text Generation

CoSyn: Detecting Implicit Hate Speech in Online Conversations Using a Context Synergized Hyperbolic Network

2023-03-02 · Sreyan Ghosh, Manan Suri, Purva Chiniya, Utkarsh Tyagi 외

The tremendous growth of social media users interacting in online conversations has led to significant growth in hate speech, affecting people from various demographics. Most of the prior works focus on detecting explici…

Towards Automatic Online Hate Speech Intervention Generation using Pretrained Language Model

2020-10-19 · Raj Ratn Pranesh, Ambesh Shekhar, Anish Kumar

Social media harbours substantial toxic and hateful conversations today. Curbing them has emerged as a critical challenge for governments and organizations globally. Prior research has primarily concentrated on the detec…

Dialogue GenerationLanguage ModelingLanguage ModellingText Generation