arXiv · 1912.09306
Tangent Space Separability in Feedforward Neural Networks
Abstract
Hierarchical neural networks are exponentially more efficient than their corresponding "shallow" counterpart with the same expressive power, but involve huge number of parameters and require tedious amounts of training. By approximating the tangent subspace, we suggest a sparse representation that enables switching to shallow networks, GradNet after a very early training stage. Our experiments show that the proposed approximation of the metric improves and sometimes even surpasses the achievable performance of the original network significantly even after a few epochs of training the original feedforward network.
Explore related subjects
Keep this discovery
Bálint Daróczy, Rita Aleksziev, András Benczúr. 2019-12-18. Tangent Space Separability in Feedforward Neural Networks. https://arxiv.org/abs/1912.09306
Cite the original work for its findings. Save a collection to share your selection of sources.