首页 > AI前沿 > Deconstructing Stereotypes: Scope-Conditioned Generation for Effective Multilingual Counterspeech

Deconstructing Stereotypes: Scope-Conditioned Generation for Effective Multilingual Counterspeech

arXiv自然语言 2026-09-15 17:36 2 阅读 查看原文

Counterspeech (CS) - direct responses that counter online Hate Speech (HS) using reasoning and alternative viewpoints - has emerged as an alternative to content removal.

Current automatic CS generation methods, however, frequently produce generic, ineffective replies that fail to target the implicit stereotypes behind HS.

To bridge this gap, we propose a novel scope-conditioned generation framework that explicitly integrates structured stereotype characteristics into Large Language Models prompts.

We validate our approach on a novel, human-curated dataset annotated in English, Italian, and Spanish.

Extensive evaluations show that stereotype-conditioned prompting substantially outperforms generic baselines across all three languages, obtaining significant gains in factuality, specificity, cogency, and effectiveness for both explicit and implicit implied stereotypes.