arXiv · 2504.16430
MAGIC: Near-Optimal Data Attribution for Deep Learning
Abstract
The goal of predictive data attribution is to estimate how adding or removing a given set of training datapoints will affect model predictions. In convex settings, this goal is straightforward (i.e., via the infinitesimal jackknife). In large-scale (non-convex) settings, however, existing methods are far less successful -- current methods' estimates often only weakly correlate with ground truth. In this work, we present a new data attribution method (MAGIC) that combines classical methods and recent advances in metadifferentiation to (nearly) optimally estimate the effect of adding or removing training data on model predictions.
Explore related subjects
Keep this discovery
Andrew Ilyas, Logan Engstrom. 2025-04-23. MAGIC: Near-Optimal Data Attribution for Deep Learning. https://arxiv.org/abs/2504.16430
Cite the original work for its findings. Save a collection to share your selection of sources.