Pith. sign in

Finite Regret and Cycles with Fixed Step-Size via Alternating Gradient Descent-Ascent

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it
abstract

Gradient descent is arguably one of the most popular online optimization methods with a wide array of applications. However, the standard implementation where agents simultaneously update their strategies yields several undesirable properties; strategies diverge away from equilibrium and regret grows over time. In this paper, we eliminate these negative properties by introducing a different implementation to obtain finite regret via arbitrary fixed step-size. We obtain this surprising property by having agents take turns when updating their strategies. In this setting, we show that an agent that uses gradient descent obtains bounded regret -- regardless of how their opponent updates their strategies. Furthermore, we show that in adversarial settings that agents' strategies are bounded and cycle when both are using the alternating gradient descent algorithm.

fields

cs.LG 1

years

2019 1

verdicts

CONDITIONAL 1

representative citing papers

Convergence of Gradient Methods on Bilinear Zero-Sum Games

cs.LG · 2019-08-15 · conditional · novelty 7.0

For bilinear zero-sum games, the paper derives necessary and sufficient convergence conditions and optimal linear rates for generalized GD, EG, OGD, and momentum methods, with Gauss-Seidel updates converging in a larger region than Jacobi updates.

citing papers explorer

Showing 1 of 1 citing paper.

  • Convergence of Gradient Methods on Bilinear Zero-Sum Games cs.LG · 2019-08-15 · conditional · none · ref 5 · internal anchor

    For bilinear zero-sum games, the paper derives necessary and sufficient convergence conditions and optimal linear rates for generalized GD, EG, OGD, and momentum methods, with Gauss-Seidel updates converging in a larger region than Jacobi updates.