Pith. sign in

Paper Citation Record · LEDGER

A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2206.05825.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2206.05825 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T19:53:25.787601Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-08T03:24:28.820322Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 114ccd43-414f-413a-be1a-983e04ccef09 · inbound

Two-Player Zero-Sum Differential Games with One-Sided Information cites this paper.

Two-Player Zero-Sum Differential Games with One-Sided Information A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-08T19:53:25.787601Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T19:53:25.787601Z digest=sha256:a119fef49ae828fec7688ca400fdae5374b05e42c0d535eee3afa8327a04233d

Observation 190c2580-1e85-4e8c-9c74-2aeb7f8369c0 · inbound

Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games cites this paper.

Incentivize without Bonus: Provably Efficient Model-based Online Multi-agent RL for Markov Games A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-07T20:38:32.878385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:38:32.878385Z digest=sha256:e5569ae04773f51fae266edb638f8f8e79c346d82a41966153b014308f83ac52

Observation 2b233b10-28b9-47d9-b0cd-ab49eb3aba8d · inbound

Multiplayer Nash Preference Optimization cites this paper.

Multiplayer Nash Preference Optimization A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T13:11:23.978594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T13:09:54.433720Z digest=sha256:56e7f6f5b23a48b856d5ec91ff383f21f184af10703544de262704b663389c5e

Observation a502d113-a250-44f7-af66-af92488bbcc4 · inbound

Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning cites this paper.

Outbidding and Outbluffing Elite Humans: Mastering Liar's Poker via Self-Play and Reinforcement Learning A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:00:34.366221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-18T00:58:51.759880Z digest=sha256:cf11c54e6b61e9f37fbb18c5b23ae06d5dd017ba3522afab7503838a8d1a4631

Observation abf03fa3-2d44-4fdf-bc52-d311d7d9fb5b · inbound

GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning cites this paper.

GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-20T07:53:10.001880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T07:50:09.840762Z digest=sha256:bde753b351c4982741d77e0c63b161b730d47cbd57eb30487cecfc08710d4c39

Observation f6e8c50c-1c80-40b1-996c-54c4e83aa27b · inbound

Which Nash Equilibrium? Solver-Dependent Selection on Zero-Sum Nash Polytopes cites this paper.

Which Nash Equilibrium? Solver-Dependent Selection on Zero-Sum Nash Polytopes A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-07-01T18:45:58.589040Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-29T01:45:47.417624Z digest=sha256:e295a9d78afbe0da4201dd40ac95f6e4d8ed450cff0e84491e07a6ecd7bb0aa1

Observation 3d3fc67c-431d-4c74-b948-c3c202956b11 · inbound

How Much Due Diligence Before You Bid? Learning in Intractable Takeover Auctions cites this paper.

How Much Due Diligence Before You Bid? Learning in Intractable Takeover Auctions A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:04:20.866971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T07:03:22.061105Z digest=sha256:6c766d8e7f28bddc52814eaadda13237269afc4d698cd8e558add2381afd5799

Observation 4f035eef-7dd1-4219-9ce5-f1e1a664c9e5 · inbound

FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games cites this paper.

FootsiesGym: A Fighting Game Benchmark for Two-Player Zero-Sum Imperfect-Information Games A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 5

Resolution
metadata mismatch
local_arxiv, observed 2026-07-08T03:24:28.821966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-07-08T03:17:09.608451Z digest=sha256:f377164ac844d14ad3dc0f0ae415bbacf0915b24f53bf9eed2f5a7dbdd88360c

Observation 2a9e4b30-0f75-48d6-aa88-50ce89117c6c · inbound

The Curvature Shadow: An Apparent Failure of Maximum-Entropy Equilibrium Selection is a Removable Artifact cites this paper.

The Curvature Shadow: An Apparent Failure of Maximum-Entropy Equilibrium Selection is a Removable Artifact A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-01T17:45:29.706313Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T17:45:29.706313Z digest=sha256:8a4260f2824210213b2e48c874aac7df2ad78332e9d97ebfe3c8ec2cd0d76dcd

Observation 32ff7f07-4263-46a4-b9a3-51803c1ebd9e · inbound

IFlowNets: Extending Generative Samplers to Learn Strategies in Incomplete Information Games cites this paper.

IFlowNets: Extending Generative Samplers to Learn Strategies in Incomplete Information Games A Unified Approach to Reinforcement Learning, Quantal Response Equilibria, and Two-Player Zero-Sum Games

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T13:37:13.048094Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T13:37:13.048094Z digest=sha256:57309f6ff7752efc1d5579ef3caad540714089abf108ff4a480395cce35dcd7e