REVIEW 1 cited by
Decentralized Multi-Agent Reinforcement Learning for Continuous-Space Stochastic Games
Not yet reviewed by Pith; the record is open.
This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.
SPECIMEN: schema-true, not a live event
T0 review · schema-true
One-sentence machine reading of the paper's core claim.
pith:XXXXXXXX · record.json · timestamp
read the original abstract
Stochastic games are a popular framework for studying multi-agent reinforcement learning (MARL). Recent advances in MARL have focused primarily on games with finitely many states. In this work, we study multi-agent learning in stochastic games with general state spaces and an information structure in which agents do not observe each other's actions. In this context, we propose a decentralized MARL algorithm and we prove the near-optimality of its policy updates. Furthermore, we study the global policy-updating dynamics for a general class of best-reply based algorithms and derive a closed-form characterization of convergence probabilities over the joint policy space.
Forward citations
Cited by 1 Pith paper
-
Equilibrium stability as a driver of cooperation among Q-learners
Q-learners with constant exploration in the repeated prisoner's dilemma spend most of their time on cooperative win-stay/lose-shift play above a boundary derived from Q-value gaps, matching simulations.
Discussion (0). Continue with ORCID to comment.