arXiv · 2605.00740
Randomized Subspace Nesterov Accelerated Gradient
Abstract
Randomized-subspace methods reduce the cost of first-order optimization by using only low-dimensional projected-gradient information, a feature that is attractive in forward-mode automatic differentiation and communication-limited settings. While Nesterov acceleration is well understood for full-gradient and coordinate-based methods, obtaining accelerated methods for general subspace sketches that use only projected-gradient information and can improve over full-dimensional Nesterov acceleration in oracle complexity is technically nontrivial. We develop randomized-subspace Nesterov accelerated gradient methods for smooth convex and smooth strongly convex optimization under matrix smoothness and generic sketch moment assumptions. The key technical ingredient is a three-sequence formulation tailored to matrix smoothness, which recovers the corresponding classical Nesterov methods in the full-dimensional case. The resulting theory establishes accelerated oracle-complexity guarantees and makes explicit how matrix smoothness and the sketch distribution enter the complexity. It also provides a unified basis for comparing sketch families and identifying when randomized-subspace acceleration improves over full-dimensional Nesterov acceleration in oracle complexity.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Gaku Omiya, Pierre-Louis Poirion, Akiko Takeda. 2026-05-01. Randomized Subspace Nesterov Accelerated Gradient. https://arxiv.org/abs/2605.00740
Cite the original work for its findings. Save a collection to share your selection of sources.