This framework integrates structured stereotype characteristics into LLM prompts to generate more specific and cogent counterspeech. Testing across English, Italian, and Spanish shows significant gains in factuality and effectiveness against both explicit and implicit stereotypes.
HOW THIS AFFECTS YOU
●
builderYou can move beyond generic moderation bots by using stereotype-conditioned prompting for more effective content responses.
●
policyThis research provides a more nuanced technical path for automating social harm mitigation via reasoning.