arXiv · 2503.02312
Go Beyond Your Means: Unlearning with Per-Sample Gradient Orthogonalization
Abstract
Machine unlearning aims to remove the influence of problematic training data after a model has been trained. The primary challenge in machine unlearning is ensuring that the process effectively removes specified data without compromising the model's overall performance on the remaining dataset. Many existing machine unlearning methods address this challenge by carefully balancing gradient ascent on the `unlearn' data with the gradient descent on a `retain' set that represents the training data. However, in many cases the training dataset is not fully available when we wish to unlearn some concepts, because models are released without their training datasets, and one may only have access to a $\textit{small part of a training set}$. Here, we propose OrthoGrad, a novel approach that mitigates interference between the unlearn set and a small retain set rather than competing ascent and descent processes. Our method projects the gradient of the unlearn set onto the subspace orthogonal to all gradients in the retain batch, effectively avoiding any gradient interference. We demonstrate the effectiveness of OrthoGrad on multiple machine unlearning benchmarks, including automatic speech recognition, outperforming competing methods.
Explore related subjects
Keep this discovery
Aviv Shamsian, Eitan Shaar, Aviv Navon, Gal Chechik, Ethan Fetaya. 2025-03-04. Go Beyond Your Means: Unlearning with Per-Sample Gradient Orthogonalization. https://arxiv.org/abs/2503.02312
Cite the original work for its findings. Save a collection to share your selection of sources.