arXiv · 2007.00389
Single Shot Structured Pruning Before Training
Abstract
We introduce a method to speed up training by 2x and inference by 3x in deep neural networks using structured pruning applied before training. Unlike previous works on pruning before training which prune individual weights, our work develops a methodology to remove entire channels and hidden units with the explicit aim of speeding up training and inference. We introduce a compute-aware scoring mechanism which enables pruning in units of sensitivity per FLOP removed, allowing even greater speed ups. Our method is fast, easy to implement, and needs just one forward/backward pass on a single batch of data to complete pruning before training begins.
Explore related subjects
Keep this discovery
Joost van Amersfoort, Milad Alizadeh, Sebastian Farquhar, Nicholas Lane, Yarin Gal. 2020-07-01. Single Shot Structured Pruning Before Training. https://arxiv.org/abs/2007.00389
Cite the original work for its findings. Save a collection to share your selection of sources.