paper-with-me

Papers

Deconstructing Stereotypes: Scope-Conditioned Generation for Effective Multilingual Counterspeech

2026-09-15 · Greta Damo, Elias Urios Alacreu, Elena Cabrio, Paolo Rosso, Serena Villata arxiv

Counterspeech (CS) - direct responses that counter online Hate Speech (HS) using reasoning and alternative viewpoints - has emerged as an alternative to content removal. Current automatic CS generation methods, however, frequently produce generic, ineffective replies that fail to target the implicit stereotypes behind HS. To bridge this gap, we propose a novel scope-conditioned generation framework that explicitly integrates structured stereotype characteristics into Large Language Models prompts. We validate our approach on a novel, human-curated dataset annotated in English, Italian, and Spanish. Extensive evaluations show that stereotype-conditioned prompting substantially outperforms generic baselines across all three languages, obtaining significant gains in factuality, specificity, cogency, and effectiveness for both explicit and implicit implied stereotypes.

📄 PDF Abstract BibTeX arXiv:2609.16906

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Queer People are People First: Deconstructing Sexual Identity Stereotypes in Large Language Models

2023-06-30 · Harnoor Dhingra, Preetiha Jayashanker, Sayali Moghe, Emma Strubell

Large Language Models (LLMs) are trained primarily on minimally processed web text, which exhibits the same wide range of social biases held by the humans who created that content. Consequently, text generated by LLMs ca…

Sentence

Investigating Gender Bias in LLM-Generated Stories via Psychological Stereotypes

2025-08-05 · Shahed Masoudian, Gustavo Escobedo, Hannah Strauss, Markus Schedl arxiv

As Large Language Models (LLMs) are increasingly used across different applications, concerns about their potential to amplify gender biases in various tasks are rising. Prior research has often probed gender bias using …

Sentence CompletionQuestion Answering

RS-Corrector: Correcting the Racial Stereotypes in Latent Diffusion Models

2023-12-08 · Yue Jiang, Yueming Lyu, Tianxiang Ma, Bo Peng 외

Recent text-conditioned image generation models have demonstrated an exceptional capacity to produce diverse and creative imagery with high visual quality. However, when pre-trained on billion-sized datasets randomly col…

Image Generation

Revisiting The Classics: A Study on Identifying and Rectifying Gender Stereotypes in Rhymes and Poems

2024-03-18 · Aditya Narayan Sankaran, Vigneshwaran Shankaran, Sampath Lonka, Rajesh Sharma

Rhymes and poems are a powerful medium for transmitting cultural norms and societal roles. However, the pervasive existence of gender stereotypes in these works perpetuates biased perceptions and limits the scope of indi…

Language ModelingLanguage ModellingLarge Language Model

Intent-conditioned and Non-toxic Counterspeech Generation using Multi-Task Instruction Tuning with RLAIF

2024-03-15 · Amey Hengle, Aswini Kumar, Sahajpreet Singh, Anil Bandhakavi 외

Counterspeech, defined as a response to mitigate online hate speech, is increasingly used as a non-censorial solution. Addressing hate speech effectively involves dispelling the stereotypes, prejudices, and biases often …

Sentence