arXiv · 2410.04577
Robustness Reprogramming for Representation Learning
Abstract
This work tackles an intriguing and fundamental open challenge in representation learning: Given a well-trained deep learning model, can it be reprogrammed to enhance its robustness against adversarial or noisy input perturbations without altering its parameters? To explore this, we revisit the core feature transformation mechanism in representation learning and propose a novel non-linear robust pattern matching technique as a robust alternative. Furthermore, we introduce three model reprogramming paradigms to offer flexible control of robustness under different efficiency requirements. Comprehensive experiments and ablation studies across diverse learning models ranging from basic linear model and MLPs to shallow and modern deep ConvNets demonstrate the effectiveness of our approaches. This work not only opens a promising and orthogonal direction for improving adversarial defenses in deep learning beyond existing methods but also provides new insights into designing more resilient AI systems with robust statistics.
Explore related subjects
Keep this discovery
Zhichao Hou, MohamadAli Torkamani, Hamid Krim, Xiaorui Liu. 2024-10-06. Robustness Reprogramming for Representation Learning. https://arxiv.org/abs/2410.04577
Cite the original work for its findings. Save a collection to share your selection of sources.