arXiv · 2011.12341
Sequential convergence of AdaGrad algorithm for smooth convex optimization
Abstract
We prove that the iterates produced by, either the scalar step size variant, or the coordinatewise variant of AdaGrad algorithm, are convergent sequences when applied to convex objective functions with Lipschitz gradient. The key insight is to remark that such AdaGrad sequences satisfy a variable metric quasi-Fej\'er monotonicity property, which allows to prove convergence.
Explore related subjects
Keep this discovery
Cheik Traoré, Edouard Pauwels. 2020-11-24. Sequential convergence of AdaGrad algorithm for smooth convex optimization. https://arxiv.org/abs/2011.12341
Cite the original work for its findings. Save a collection to share your selection of sources.