REVIEW 1 cited by
Stochastic Multiplicative Weights Updates in Zero-Sum Games
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
We study agents competing against each other in a repeated network zero-sum game while applying the multiplicative weights update (MWU) algorithm with fixed learning rates. In our implementation, agents select their strategies probabilistically in each iteration and update their weights/strategies using the realized vector payoff of all strategies, i.e., stochastic MWU with full information. We show that the system results in an irreducible Markov chain where agent strategies diverge from the set of Nash equilibria. Further, we show that agents will play pure strategies with probability 1 in the limit.
Forward citations
Cited by 1 Pith paper
-
Learnable Mixed Nash Equilibria are Collectively Rational
A mixed Nash equilibrium that is locally uniformly stable under uncoupled learning dynamics must be weakly Pareto optimal, and uniform stability controls last-iterate convergence of smoothed best-response dynamics.
Discussion (0). Continue with ORCID to comment.