Pith. sign in

REVIEW 5 cited by

Policy Space Response Oracles: A Survey

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2403.02227 v2 pith:ZPXR5RBN submitted 2024-03-04 cs.GT cs.AIcs.MA

classification cs.GTcs.AIcs.MA
keywords psrostrategiessurveygamelargeoraclespolicyprovides
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

Game theory provides a mathematical way to study the interaction between multiple decision makers. However, classical game-theoretic analysis is limited in scalability due to the large number of strategies, precluding direct application to more complex scenarios. This survey provides a comprehensive overview of a framework for large games, known as Policy Space Response Oracles (PSRO), which holds promise to improve scalability by focusing attention on sufficient subsets of strategies. We first motivate PSRO and provide historical context. We then focus on the strategy exploration problem for PSRO: the challenge of assembling effective subsets of strategies that still represent the original game well with minimum computational cost. We survey current research directions for enhancing the efficiency of PSRO, and explore the applications of PSRO across various domains. We conclude by discussing open questions and future research.

Discussion (0). Sign in to comment.

Forward citations

Cited by 5 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score.

  1. Beyond Bayesian Nash: Learning Minimax-Regret Equilibria for Adversarial Team Games under Asymmetric Information

    cs.GT 2026-07 conditional novelty 6.5 of 10

    PR-MRE hedges regret over high-confidence type subsets and, via PRMRE-PSRO, yields more distribution-shift-robust team policies than BNE on graph Capture-the-Flag.

  2. Evolving in the Agent Jungle via History-Informed Opponent Awareness

    cs.AI 2026-08 conditional novelty 6.0 of 10

    OASE filters LLM agent skill revisions through paired tests against historical opponent snapshots, yielding lower equilibrium distance and fewer accepted edits in auctions and Cournot games.

  3. Offline Nash Solvers Meet Online Tree Search in Multi-Agent Games on Graphs

    cs.GT 2026-07 conditional novelty 6.0 of 10

    Primitive-Guided Tree Search combines offline exact Nash solutions of 1v1/2v1 subgames with online SM-MCTS to produce coordinated multi-agent pursuit policies that outperform PSRO and neural MCTS baselines.

  4. CyGym: A Simulation-Based Game-Theoretic Analysis Framework for Cybersecurity

    cs.CR 2025-06 conditional novelty 6.0 of 10

    CyGym provides a Gym-based cyber simulation with a POSG formalization, zero-day modeling, and a PSRO-style solver, evaluated on a Volt Typhoon scenario.

  5. DyGRO-VLA: Cross-Task Scaling of Vision-Language-Action Models via Dynamic Grouped Residual Optimization

    cs.RO 2026-05 unverdicted novelty 5.0 of 10

    DyGRO-VLA is a two-stage optimization framework for cross-task scaling of Vision-Language-Action models via dynamic grouped residual optimization in RL.

Pith tools