arXiv · 2305.12394
Pruning Pre-trained Language Models with Principled Importance and Self-regularization
Abstract
Iterative pruning is one of the most effective compression methods for pre-trained language models. We discovered that finding the optimal pruning decision is an equality-constrained 0-1 Integer Linear Programming problem. The solution to this optimization problem leads to a principled importance criterion which we use to rank parameters during iterative model pruning. To mitigate the poor generalization at high sparsity levels, we propose a self-regularization scheme where model prediction is regularized by the latest checkpoint with increasing sparsity throughout pruning. Our experiments on natural language understanding, question-answering, named entity recognition, and data-to-text generation with various Transformer-based PLMs show the effectiveness of the approach at various sparsity levels.
Explore related subjects
Keep this discovery
Siyu Ren, Kenny Q. Zhu. 2023-05-21. Pruning Pre-trained Language Models with Principled Importance and Self-regularization. https://arxiv.org/abs/2305.12394
Cite the original work for its findings. Save a collection to share your selection of sources.