arXiv · 2508.20773
Unleashing Uncertainty: Efficient Machine Unlearning for Generative AI
Abstract
We introduce SAFEMax, a novel method for Machine Unlearning in diffusion models. Grounded in information-theoretic principles, SAFEMax maximizes the entropy in generated images, causing the model to generate Gaussian noise when conditioned on impermissible classes by ultimately halting its denoising process. Also, our method controls the balance between forgetting and retention by selectively focusing on the early diffusion steps, where class-specific information is prominent. Our results demonstrate the effectiveness of SAFEMax and highlight its substantial efficiency gains over state-of-the-art methods.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Christoforos N. Spartalis, Theodoros Semertzidis, Petros Daras, Efstratios Gavves. 2025-08-28. Unleashing Uncertainty: Efficient Machine Unlearning for Generative AI. https://arxiv.org/abs/2508.20773
Cite the original work for its findings. Save a collection to share your selection of sources.