arXiv · 2402.05025
Strong convexity-guided hyper-parameter optimization for flatter losses
Abstract
We propose a novel white-box approach to hyper-parameter optimization. Motivated by recent work establishing a relationship between flat minima and generalization, we first establish a relationship between the strong convexity of the loss and its flatness. Based on this, we seek to find hyper-parameter configurations that improve flatness by minimizing the strong convexity of the loss. By using the structure of the underlying neural network, we derive closed-form equations to approximate the strong convexity parameter, and attempt to find hyper-parameters that minimize it in a randomized fashion. Through experiments on 14 classification datasets, we show that our method achieves strong performance at a fraction of the runtime.
Explore related subjects
Keep this discovery
Rahul Yedida, Snehanshu Saha. 2024-02-07. Strong convexity-guided hyper-parameter optimization for flatter losses. https://arxiv.org/abs/2402.05025
Cite the original work for its findings. Save a collection to share your selection of sources.