arXiv · 2312.06957
Online Saddle Point Problem and Online Convex-Concave Optimization
Abstract
Centered around solving the Online Saddle Point problem, this paper introduces the Online Convex-Concave Optimization (OCCO) framework, which involves a sequence of two-player time-varying convex-concave games. We propose the generalized duality gap (Dual-Gap) as the performance metric and establish the parallel relationship between OCCO with Dual-Gap and Online Convex Optimization (OCO) with regret. To demonstrate the natural extension of OCCO from OCO, we develop two algorithms, the implicit online mirror descent-ascent and its optimistic variant. Analysis reveals that their duality gaps share similar expression forms with the corresponding dynamic regrets arising from implicit updates in OCO. Empirical results further substantiate the effectiveness of our algorithms. Simultaneously, we unveil that the dynamic Nash equilibrium regret, which was initially introduced in a recent paper, has inherent defects.
Explore related subjects
Keep this discovery
Qing-xin Meng, Jian-wei Liu. 2023-12-12. Online Saddle Point Problem and Online Convex-Concave Optimization. https://arxiv.org/abs/2312.06957
Cite the original work for its findings. Save a collection to share your selection of sources.