arXiv · 2106.06477
Avoiding local minima in multilayer network optimization by incremental training
Abstract
Training a large multilayer neural network can present many difficulties due to the large number of useless stationary points. These points usually attract the minimization algorithm used during the training phase, which therefore results inefficient. Extending some results proposed in literature for shallow networks, we propose the mathematical characterization of a class of such stationary points that arise in deep neural networks training. Availing such a description, we are able to define an incremental training algorithm that avoids getting stuck in the region of attraction of these undesirable stationary points.
Explore related subjects
Keep this discovery
Alberto De Santis, Giampaolo Liuzzi, Stefano Lucidi, Edoardo Maria Tronci. 2021-06-11. Avoiding local minima in multilayer network optimization by incremental training. https://arxiv.org/abs/2106.06477
Cite the original work for its findings. Save a collection to share your selection of sources.