arXiv · 2209.14558
Computational Complexity of Sub-Linear Convergent Algorithms
Abstract
Optimizing machine learning algorithms that are used to solve the objective function has been of great interest. Several approaches to optimize common algorithms, such as gradient descent and stochastic gradient descent, were explored. One of these approaches is reducing the gradient variance through adaptive sampling to solve large-scale optimization's empirical risk minimization (ERM) problems. In this paper, we will explore how starting with a small sample and then geometrically increasing it and using the solution of the previous sample ERM to compute the new ERM. This will solve ERM problems with first-order optimization algorithms of sublinear convergence but with lower computational complexity. This paper starts with theoretical proof of the approach, followed by two experiments comparing the gradient descent with the adaptive sampling of the gradient descent and ADAM with adaptive sampling ADAM on different datasets.
Explore related subjects
Keep this discovery
Hilal AlQuabeh, Farha AlBreiki, Dilshod Azizov. 2022-09-29. Computational Complexity of Sub-Linear Convergent Algorithms. https://arxiv.org/abs/2209.14558
Cite the original work for its findings. Save a collection to share your selection of sources.