arXiv · 2112.09151
TAFIM: Targeted Adversarial Attacks against Facial Image Manipulations
Abstract
Face manipulation methods can be misused to affect an individual's privacy or to spread disinformation. To this end, we introduce a novel data-driven approach that produces image-specific perturbations which are embedded in the original images. The key idea is that these protected images prevent face manipulation by causing the manipulation model to produce a predefined manipulation target (uniformly colored output image in our case) instead of the actual manipulation. In addition, we propose to leverage differentiable compression approximation, hence making generated perturbations robust to common image compression. In order to prevent against multiple manipulation methods simultaneously, we further propose a novel attention-based fusion of manipulation-specific perturbations. Compared to traditional adversarial attacks that optimize noise patterns for each image individually, our generalized model only needs a single forward pass, thus running orders of magnitude faster and allowing for easy integration in image processing stacks, even on resource-constrained devices like smartphones.
Explore related subjects
Keep this discovery
Shivangi Aneja, Lev Markhasin, Matthias Niessner. 2021-12-16. TAFIM: Targeted Adversarial Attacks against Facial Image Manipulations. https://arxiv.org/abs/2112.09151
Cite the original work for its findings. Save a collection to share your selection of sources.