INCLUDE Benchmark Quantifies Cross-Lingual Safety Gaps in LLMs
August 20, 2026
The INCLUDE benchmark uses 2,604 prompts across six languages to measure socio-cultural biases in multilingual models. It reveals significant safety alignment failures in non-English languages, particularly within the Indian linguistic context.
HOW THIS AFFECTS YOU
●
researcherYou can use this benchmark to evaluate how safety alignment transfers across linguistically diverse datasets.
●
policyThis highlights the regulatory and ethical risks of deploying English-centric safety filters globally.