mu2-Bench for Multilingual Machine Unlearning Evaluation
September 21, 2026
mu2-Bench evaluates how effectively multilingual LLMs remove undesired or private information across different languages. The benchmark tests whether unlearning target knowledge in one language successfully prevents its cross-linguistic propagation in others.
HOW THIS AFFECTS YOU
●
researcherYou can use this to study the cross-linguistic spread of memorized data during unlearning.
●
policyThis helps you assess the robustness of data privacy and safety removal across global language models.