arXiv · 2410.03640
Real-World Benchmarks Make Membership Inference Attacks Fail on Diffusion Models
Abstract
Membership inference attacks (MIAs) on diffusion models have emerged as potential evidence of unauthorized data usage in training pre-trained diffusion models. These attacks aim to detect the presence of specific images in training datasets of diffusion models. Our study delves into the evaluation of state-of-the-art MIAs on diffusion models and reveals critical flaws and overly optimistic performance estimates in existing MIA evaluation. We introduce CopyMark, a more realistic MIA benchmark that distinguishes itself through the support for pre-trained diffusion models, unbiased datasets, and fair evaluation pipelines. Through extensive experiments, we demonstrate that the effectiveness of current MIA methods significantly degrades under these more practical conditions. Based on our results, we alert that MIA, in its current state, is not a reliable approach for identifying unauthorized data usage in pre-trained diffusion models. To the best of our knowledge, we are the first to discover the performance overestimation of MIAs on diffusion models and present a unified benchmark for more realistic evaluation. Our code is available on GitHub: \url{https://github.com/caradryanl/CopyMark}.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Chumeng Liang, Jiaxuan You. 2024-10-04. Real-World Benchmarks Make Membership Inference Attacks Fail on Diffusion Models. https://arxiv.org/abs/2410.03640
Cite the original work for its findings. Save a collection to share your selection of sources.