paper-with-me

홈 › Papers

Understanding Counterspeech for Online Harm Mitigation

2023-07-01 · Yi-Ling Chung, Gavin Abercrombie, Florence Enock, Jonathan Bright, Verena Rieser

Counterspeech offers direct rebuttals to hateful speech by challenging perpetrators of hate and showing support to targets of abuse. It provides a promising alternative to more contentious measures, such as content moderation and deplatforming, by contributing a greater amount of positive online speech rather than attempting to mitigate harmful content through removal. Advances in the development of large language models mean that the process of producing counterspeech could be made more efficient by automating its generation, which would enable large-scale online campaigns. However, we currently lack a systematic understanding of several important factors relating to the efficacy of counterspeech for hate mitigation, such as which types of counterspeech are most effective, what are the optimal conditions for implementation, and which specific effects of hate it can best ameliorate. This paper aims to fill this gap by systematically reviewing counterspeech research in the social sciences and comparing methodologies and findings with computer science efforts in automatic counterspeech generation. By taking this multi-disciplinary view, we identify promising future directions in both fields.

📄 PDF Abstract BibTeX arXiv:2307.04761

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

NLP Systems That Can't Tell Use from Mention Censor Counterspeech, but Teaching the Distinction Helps

2024-04-02 · Kristina Gligoric, Myra Cheng, Lucia Zheng, Esin Durmus 외

The use of words to convey speaker's intent is traditionally distinguished from the `mention' of words for quoting what someone said, or pointing out properties of a word. Here we show that computationally modeling this …

Hate Speech DetectionMisinformation

CounterQuill: Investigating the Potential of Human-AI Collaboration in Online Counterspeech Writing

2024-10-03 · Xiaohan Ding, Kaike Ping, Uma Sushmitha Gunturi, Buse Carik 외

Online hate speech has become increasingly prevalent on social media, causing harm to individuals and society. While automated content moderation has received considerable attention, user-driven counterspeech remains a l…

Is Safer Better? The Impact of Guardrails on the Argumentative Strength of LLMs in Hate Speech Countering

2024-10-04 · Helena Bonaldi, Greta Damo, Nicolás Benjamín Ocampo, Elena Cabrio 외

The potential effectiveness of counterspeech as a hate speech mitigation strategy is attracting increasing interest in the NLG research community, particularly towards the task of automatically producing it. However, aut…

Speaking at the Right Level: Literacy-Controlled Counterspeech Generation with RAG-RL

2025-09-01 · Xiaoying Song, Anirban Saha Anik, Dibakar Barua, Pengcheng Luo 외 arxiv

Health misinformation spreading online poses a significant threat to public health. Researchers have explored methods for automatically generating counterspeech to health misinformation as a mitigation strategy. Existing…

Reinforcement Learning

Echoes of Discord: Forecasting Hater Reactions to Counterspeech

2025-01-27 · Xiaoying Song, Sharon Lisseth Perez, Xinchen Yu, Eduardo Blanco 외

Hate speech (HS) erodes the inclusiveness of online users and propagates negativity and division. Counterspeech has been recognized as a way to mitigate the harmful consequences. While some research has investigated the …