arXiv · 2608.20712
Privacy-Preserving Object Detection for Vision Transformer-Based Models
Abstract
We propose a novel object detection method that enables us to protect sensitive visual information of test images. Previous studies considering visual information protection focus on image classification tasks. This paper proposes an object detection method using perceptual encryption for the first time. The proposed method can achieve almost the same accuracy as that of models without any protection by utilizing the embedding structure of the Vision Transformer (ViT) and a domain adaptation technique with keys. In experiments, the effectiveness of the proposed method is verified in terms of accuracy and visual protection under the use of ViTdet, which is a ViT-based object detection model.
Explore related subjects
Keep this discovery
Homare Sueyoshi, Kiyoshi Nishikawa, Hitoshi Kiya. 2026-08-21. Privacy-Preserving Object Detection for Vision Transformer-Based Models. https://arxiv.org/abs/2608.20712
Cite the original work for its findings. Save a collection to share your selection of sources.