arXiv · 2502.11619
Membership Inference Attacks for Face Images Against Fine-Tuned Latent Diffusion Models
Abstract
The rise of generative image models leads to privacy concerns when it comes to the huge datasets used to train such models. This paper investigates the possibility of inferring if a set of face images was used for fine-tuning a Latent Diffusion Model (LDM). A Membership Inference Attack (MIA) method is presented for this task. Using generated auxiliary data for the training of the attack model leads to significantly better performance, and so does the use of watermarks. The guidance scale used for inference was found to have a significant influence. If a LDM is fine-tuned for long enough, the text prompt used for inference has no significant influence. The proposed MIA is found to be viable in a realistic black-box setup against LDMs fine-tuned on face-images.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Lauritz Christian Holme, Anton Mosquera Storgaard, Siavash Arjomand Bigdeli. 2025-02-17. Membership Inference Attacks for Face Images Against Fine-Tuned Latent Diffusion Models. https://arxiv.org/abs/2502.11619
Cite the original work for its findings. Save a collection to share your selection of sources.