arXiv · 2412.06070
Stochastic Gradient Descent Revisited
Abstract
Stochastic gradient descent (SGD) has been a go-to algorithm for nonconvex stochastic optimization problems arising in machine learning. Its theory however often requires a strong framework to guarantee convergence properties. We hereby present a full scope convergence study of biased nonconvex SGD, including weak convergence, function-value convergence and global convergence, and also provide subsequent convergence rates and complexities, all under relatively mild conditions in comparison with literature.
Explore related subjects
Keep this discovery
Azar Louzi. 2024-12-08. Stochastic Gradient Descent Revisited. https://arxiv.org/abs/2412.06070
Cite the original work for its findings. Save a collection to share your selection of sources.