Pith. sign in

REVIEW 1 cited by

Multi-agent online learning in time-varying games

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1809.03066 v3 pith:PVS73NEU submitted 2018-09-10 cs.GT cs.LGmath.OC

classification cs.GTcs.LGmath.OC
keywords gamesequilibriumfeedbacklearningmonotonemulti-agentonlinesequence
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We examine the long-run behavior of multi-agent online learning in games that evolve over time. Specifically, we focus on a wide class of policies based on mirror descent, and we show that the induced sequence of play (a) converges to Nash equilibrium in time-varying games that stabilize in the long run to a strictly monotone limit; and (b) it stays asymptotically close to the evolving equilibrium of the sequence of stage games (assuming they are strongly monotone). Our results apply to both gradient-based and payoff-based feedback - i.e., the "bandit feedback" case where players only get to observe the payoffs of their chosen actions.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Prediction-Aware Learning in Multi-Agent Systems

    cs.GT 2025-01 accept novelty 6.0 of 10

    A contextual optimistic multiplicative weights algorithm (POMWU) achieves static-game regret, equilibrium convergence, and social welfare guarantees in time-varying games when players can predict the changing state of...

Pith tools