arXiv · 2609.20945
$μ^2$-Bench: A Multilingual Machine Unlearning Benchmark
Abstract
Undesired information such as harmful content and private data propagates through Multilingual Large Language Models (LLMs) via direct training and indirect cross-linguistic spread. Multilingual Machine Unlearning (MMU) aims to remove such information, yet its evaluation remains underexplored, leaving unclear whether unlearning truly eliminates target knowledge across all languages. To bridge this gap, we introduce $μ^2$-Bench, an MMU benchmark that simulates the full pipeline of memorization, unlearning, and evaluation across diverse languages. It 1) spans a broad set of languages, 2) evaluates on both training and hold-out languages, and 3) assesses knowledge as dispersed across multiple languages. We show that successful MMU requires methods that reflect multilingual characteristics, and conduct analysis to provide deeper insights into MMU.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Kyomin Hwang, Hyeonjin Kim, Hyunho Lee, Yearim Kim, Yeji Song, Nojun Kwak. 2026-09-17. $μ^2$-Bench: A Multilingual Machine Unlearning Benchmark. https://arxiv.org/abs/2609.20945
Cite the original work for its findings. Save a collection to share your selection of sources.