BabelSafe is a policy-grounded safety benchmark covering 13 languages, using regional regulatory documents to guide data generation. It evaluates LLMs against jurisdiction-specific rules and cultural contexts rather than relying on generic risk taxonomies.
HOW THIS AFFECTS YOU
●
researcherYou can use this to test the cross-linguistic safety and cultural alignment of your models.
●
policyThis provides a concrete framework for auditing LLMs against localized legal and regulatory standards.