paper-with-me

홈 › Papers

Assisted Counterspeech Writing at the Crossroads of Hate Speech and Misinformation

2026-05-21 · Genoveffa Martone, Helena Bonaldi, Marco Guerini arxiv

Hate speech and misinformation frequently co-occur online, amplifying prejudice and polarization. Given their scale, using Large Language Models (LLMs) to assist expert counterspeech (CS) writing has gained interest, yet prior work has addressed these phenomena separately. We bridge this gap by studying CS generation in contexts where both hate and misinformation co-occur. We test three knowledge-driven generation strategies: first we prompt an LLM with fact-checkers' guidelines and fact-checking articles; secondly, with NGOs' guidelines and reports; thirdly, we create a mixed strategy that combines guidelines and documents from both. 23 experts revise the generated CS, which are assessed via human and automatic metrics. While LLMs produce adequate CS in 40% of cases, expert edits substantially improve naturalness, exhaustiveness, and adherence to guidelines. Based on the post-edited CS, the mixed strategy proves to be the most effective in crowdsourcing evaluation, pairing strong factual correction with stereotype mitigation and empathetic engagement. We release a dataset of hateful and misinformed claims with expert-verified CS and supporting knowledge.

📄 PDF Abstract BibTeX arXiv:2605.22435

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

CounterQuill: Investigating the Potential of Human-AI Collaboration in Online Counterspeech Writing

2024-10-03 · Xiaohan Ding, Kaike Ping, Uma Sushmitha Gunturi, Buse Carik 외

Online hate speech has become increasingly prevalent on social media, causing harm to individuals and society. While automated content moderation has received considerable attention, user-driven counterspeech remains a l…

CrowdCounter: A benchmark type-specific multi-target counterspeech dataset

2024-10-02 · Punyajoy Saha, Abhilash Datta, Abhik Jana, Animesh Mukherjee

Counterspeech presents a viable alternative to banning or suspending users for hate speech while upholding freedom of expression. However, writing effective counterspeech is challenging for moderators/users. Hence, devel…

Diversity

Echoes of Discord: Forecasting Hater Reactions to Counterspeech

2025-01-27 · Xiaoying Song, Sharon Lisseth Perez, Xinchen Yu, Eduardo Blanco 외

Hate speech (HS) erodes the inclusiveness of online users and propagates negativity and division. Counterspeech has been recognized as a way to mitigate the harmful consequences. While some research has investigated the …

Racism is a Virus: Anti-Asian Hate and Counterspeech in Social Media during the COVID-19 Crisis

2020-05-25 · Bing He, Caleb Ziems, Sandeep Soni, Naren Ramakrishnan 외

The spread of COVID-19 has sparked racism and hate on social media targeted towards Asian communities. However, little is known about how racial hate spreads during a pandemic and the role of counterspeech in mitigating …

Beyond Denouncing Hate: Strategies for Countering Implied Biases and Stereotypes in Language

2023-10-31 · Jimin Mun, Emily Allaway, Akhila Yerukola, Laura Vianna 외

Counterspeech, i.e., responses to counteract potential harms of hateful speech, has become an increasingly popular solution to address online hate speech without censorship. However, properly countering hateful language …

Philosophy