arXiv · 2512.24999
Basic Inequalities for First-Order Optimization with Applications to Statistical Risk Analysis
Abstract
We introduce \textit{basic inequalities} for first-order iterative optimization algorithms, forming a simple and versatile framework that connects implicit and explicit regularization. While related inequalities appear in the literature, we isolate and highlight a specific form and develop it as a well-rounded tool for statistical analysis. Let $f$ denote the objective function to be optimized. Given a first-order iterative algorithm initialized at $\theta_0$ with current iterate $\theta_T$, the basic inequality upper bounds $f(\theta_T)-f(z)$ for any reference point $z$ in terms of the accumulated step sizes and the distances between $\theta_0$, $\theta_T$, and $z$. The bound translates the number of iterations into an effective regularization coefficient in the loss function. We demonstrate this framework through analyses of training dynamics and prediction risk bounds. In addition to revisiting and refining known results on gradient descent, we provide new results for mirror descent with Bregman divergence projection, for generalized linear models trained by gradient descent and exponentiated gradient descent, and for randomized predictors. We illustrate and supplement these theoretical findings with experiments on generalized linear models.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Seunghoon Paik, Kangjie Zhou, Matus Telgarsky, Ryan J. Tibshirani. 2025-12-31. Basic Inequalities for First-Order Optimization with Applications to Statistical Risk Analysis. https://arxiv.org/abs/2512.24999
Cite the original work for its findings. Save a collection to share your selection of sources.