arXiv · 2406.13888
Open Problem: Anytime Convergence Rate of Gradient Descent
Abstract
Recent results show that vanilla gradient descent can be accelerated for smooth convex objectives, merely by changing the stepsize sequence. We show that this can lead to surprisingly large errors indefinitely, and therefore ask: Is there any stepsize schedule for gradient descent that accelerates the classic $\mathcal{O}(1/T)$ convergence rate, at \emph{any} stopping time $T$?
Explore related subjects
Keep this discovery
Guy Kornowski, Ohad Shamir. 2024-06-19. Open Problem: Anytime Convergence Rate of Gradient Descent. https://arxiv.org/abs/2406.13888
Cite the original work for its findings. Save a collection to share your selection of sources.