paper-with-me

Papers

Parsimonious Argument Annotations for Hate Speech Counter-narratives

2022-08-01 · Damian A. Furman, Pablo Torres, Jose A. Rodriguez, Lautaro Martinez, Laura Alonso Alemany, Diego Letzen, Maria Vanina Martinez

We present an enrichment of the Hateval corpus of hate speech tweets (Basile et. al 2019) aimed to facilitate automated counter-narrative generation. Comparably to previous work (Chung et. al. 2019), manually written counter-narratives are associated to tweets. However, this information alone seems insufficient to obtain satisfactory language models for counter-narrative generation. That is why we have also annotated tweets with argumentative information based on Wagemanns (2016), that we believe can help in building convincing and effective counter-narratives for hate speech against particular groups. We discuss adequacies and difficulties of this annotation process and present several baselines for automatic detection of the annotated elements. Preliminary results show that automatic annotators perform close to human annotators to detect some aspects of argumentation, while others only reach low or moderate level of inter-annotator agreement.

📄 PDF Abstract BibTeX arXiv:2208.01099

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Is Safer Better? The Impact of Guardrails on the Argumentative Strength of LLMs in Hate Speech Countering

2024-10-04 · Helena Bonaldi, Greta Damo, Nicolás Benjamín Ocampo, Elena Cabrio 외

The potential effectiveness of counterspeech as a hate speech mitigation strategy is attracting increasing interest in the NLG research community, particularly towards the task of automatically producing it. However, aut…

Leveraging Argument Structure to Predict Content Hatefulness

2026-05-04 · Nicolás Benjamín Ocampo, Davide Ceolin arxiv

Information disorder is a challenging phenomenon that affects society at large. This phenomenon entails the diffusion of misleading, misinforming, and hateful content online. In different contexts, one aspect of the prob…

STATE ToxiCN: A Benchmark for Span-level Target-Aware Toxicity Extraction in Chinese Hate Speech Detection

2025-01-26 · Zewen Bai, Shengdi Yin, Junyu Lu, Jingjie Zeng 외

The proliferation of hate speech has caused significant harm to society. The intensity and directionality of hate are closely tied to the target and argument it is associated with. However, research on hate speech detect…

Hate Speech Detection

Consolidating Strategies for Countering Hate Speech Using Persuasive Dialogues

2024-01-15 · Sougata Saha, Rohini Srihari

Hateful comments are prevalent on social media platforms. Although tools for automatically detecting, flagging, and blocking such false, offensive, and harmful content online have lately matured, such reactive and brute …

BlockingResponse Generation

Counter with Evidence! A Multi-Agent Memory Efficient Reasoning Framework for Hate Category Informed Counterspeech Generation

2026-08-24 · Sujoy Nath, Aswini Kumar, Tanmoy Chakraborty arxiv

Counterspeech effectively neutralizes the impact of online hate. Although prior work explores automated counterspeech generation, it largely emphasizes stylistic control while treating hate speech as homogeneous, overloo…