Pith. sign in

REVIEW 1 cited by

Value Function Approximation in Zero-Sum Markov Games

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 1301.0580 v1 pith:LIIBYJE4 submitted 2012-12-12 cs.AI

classification cs.AI
keywords markovgamesapproximationfunctionvalueproblemboundsgeneralization
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

This paper investigates value function approximation in the context of zero-sum Markov games, which can be viewed as a generalization of the Markov decision process (MDP) framework to the two-agent case. We generalize error bounds from MDPs to Markov games and describe generalizations of reinforcement learning algorithms to Markov games. We present a generalization of the optimal stopping problem to a two-player simultaneous move Markov game. For this special problem, we provide stronger bounds and can guarantee convergence for LSTD and temporal difference learning with linear value function approximation. We demonstrate the viability of value function approximation for Markov games by using the Least squares policy iteration (LSPI) algorithm to learn good policies for a soccer domain and a flow control problem.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. A Multi-Resolution Dynamic Game Framework for Cross-Echelon Decision-Making in Cyber Warfare

    cs.CR 2025-07 conditional novelty 5.0 of 10

    A two-resolution game framework lets cyber defenders zoom between tactical game trees and strategic Markov game states to refine defense plans.

Pith tools