arXiv · 2105.14797
RED : Looking for Redundancies for Data-Free Structured Compression of Deep Neural Networks
Abstract
Deep Neural Networks (DNNs) are ubiquitous in today's computer vision land-scape, despite involving considerable computational costs. The mainstream approaches for runtime acceleration consist in pruning connections (unstructured pruning) or, better, filters (structured pruning), both often requiring data to re-train the model. In this paper, we present RED, a data-free structured, unified approach to tackle structured pruning. First, we propose a novel adaptive hashing of the scalar DNN weight distribution densities to increase the number of identical neurons represented by their weight vectors. Second, we prune the network by merging redundant neurons based on their relative similarities, as defined by their distance. Third, we propose a novel uneven depthwise separation technique to further prune convolutional layers. We demonstrate through a large variety of benchmarks that RED largely outperforms other data-free pruning methods, often reaching performance similar to unconstrained, data-driven methods.
Explore related subjects
Keep this discovery
Edouard Yvinec, Arnaud Dapogny, Matthieu Cord, Kevin Bailly. 2021-05-31. RED : Looking for Redundancies for Data-Free Structured Compression of Deep Neural Networks. https://arxiv.org/abs/2105.14797
Cite the original work for its findings. Save a collection to share your selection of sources.