Pith. sign in

Paper Citation Record · LEDGER

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games

As of 7 August 2026, this Paper Citation Record lists 63 of 63 outbound references and 0 inbound Pith citation observations for arXiv:2506.05894.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.05894 v1

Coverage vector

measured 63 of 63 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T10:25:56.452058Z

measured 63 of 63 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

63 of 63 outbound references displayed

  • verified exact2
  • verified fuzzy48
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 785a7eff-dc7e-424c-afed-7efebc62578f · outbound

This paper cites Value iteration algorithm for mean- field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Value iteration algorithm for mean- field games

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.559605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:49.554219Z digest=sha256:e1c4bb6c87310d9e61c7e6ccc1584a38414352176751d4c048e4d424819eb6e2

Observation 7db36181-eddf-4575-8802-3f5cdfca4e99 · outbound

This paper cites Unified reinforcement q-learning for mean field game and control problems.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unified reinforcement q-learning for mean field game and control problems

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.551580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:49.639583Z digest=sha256:85e181af217d2026b23ceb3cafbfd9d9deadb890af043b409943e040812b3615

Observation efcfd009-787b-483b-b5fe-2e6069da6fb5 · outbound

This paper cites Stochastic graphon games: II.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic graphon games: II

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.543561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:49.790981Z digest=sha256:b4138e90fcd7569d66b283393d140d41667db24e2f531579b1a02b8fddb7f510

Observation a2c3fda2-815f-492e-a65d-f853df8b6ba4 · outbound

This paper cites Berahas, Liyuan Cao, Krzysztof Choromanski, and Katya Scheinberg.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Berahas, Liyuan Cao, Krzysztof Choromanski, and Katya Scheinberg

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.535643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:49.928649Z digest=sha256:1eb9a7293a44bad362068d5b17888eae8b6e6490b24a8bf76826e318e1c39ef8

Observation a0e11cf2-e17c-418f-8c40-8cd5655da34e · outbound

This paper cites An Lp theory of sparse graph convergence i: Limits, sparse random graph models, and power law distributions.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games An Lp theory of sparse graph convergence i: Limits, sparse random graph models, and power law distributions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.527738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:50.035037Z digest=sha256:a54c834a0ce6bcf50190e37b7c4544724ac3f5e2bfa92adc2ee179b9159a7392

Observation 123f7447-30c8-42d1-97e2-3bb0285ef568 · outbound

This paper cites Chayes, Henry Cohn, and Yufei Zhao.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Chayes, Henry Cohn, and Yufei Zhao

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.519563Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:50.119793Z digest=sha256:6b09e9d74c34562e852b99ce258498c5a7a47060fc15772a89dc268868c298f2

Observation 3542a397-2415-4f04-afee-fc70ea509587 · outbound

This paper cites Functional Analysis, Sobolev Spaces and Partial Differential Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Functional Analysis, Sobolev Spaces and Partial Differential Equations

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.511383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:50.175222Z digest=sha256:46ceae3f99064cb26451ca9a6e2ae263758ddbc7003a33ade5393825d1976ac4

Observation 268cc0c8-0e71-4c4f-8d2b-eff23765fff7 · outbound

This paper cites Policy Gradient-based Algorithms for Continuous-time Linear Quadratic Control.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy Gradient-based Algorithms for Continuous-time Linear Quadratic Control

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:50.242646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:50.242646Z digest=sha256:5608a31279a06c9826aed119155a0cfce3e3b919930b1131051d119f9bea3243

Observation 7175448f-8be3-4a07-9c46-f35e44628ed0 · outbound

This paper cites Caines and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines and Minyi Huang

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.503030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:50.336317Z digest=sha256:0ad44a052b11f6de01cad6c710cc3f62c2a7225e93c0833b3f6a7517ff40dc0c

Observation ebf7009f-a1af-48e6-a049-0a496c937c73 · outbound

This paper cites Financial Mathematics.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Financial Mathematics

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.495330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:50.443526Z digest=sha256:28d14232b7d77b87e93301cdb0a31f65a536d5f65c87d107988c6474f89dbb6c

Observation 07a7c107-9089-4b83-9fbb-b66acd672b52 · outbound

This paper cites Cooney, Christy V.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Cooney, Christy V

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.487337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:50.539531Z digest=sha256:72829e06b04b251a8e1f9ed61c4e983e67110eb3237d3c7e8a427fc3b48ee849

Observation e669e92f-5b9c-46d4-bd91-2aff43417c13 · outbound

This paper cites Deep learning for mean field games and mean field control with applications to finance.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Deep learning for mean field games and mean field control with applications to finance

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.479292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:50.619486Z digest=sha256:647db64691a1d0cfb4d249e687ba05434ce0c2ee97f28ce0e6538e7e31dda765

Observation eec5f744-7b5d-4f39-9d94-85ca914afc08 · outbound

This paper cites Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:50.732625Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:50.732625Z digest=sha256:d368158791ec313406f2fbddf89a1c5e101b7dff6a1d2f3ce4045d87da0991a3

Observation 1004d101-8bb8-45d1-be4c-5d1933abb5e1 · outbound

This paper cites Approximately solving mean field games via entropy-regularized deep reinforcement learning.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Approximately solving mean field games via entropy-regularized deep reinforcement learning

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.471147Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:50.868633Z digest=sha256:6f4c02fb561bfe708fec6879ab0de436e1b616ace0dcc244b4e6079d252ecf12

Observation 78277913-cbb5-471b-b07c-d73702318517 · outbound

This paper cites Learning graphon mean field games and approximate Nash equilibria.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning graphon mean field games and approximate Nash equilibria

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.463392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:51.023614Z digest=sha256:f76f057354dbe1828e6ac0d433ec31040a2b0b96ad4e602940defe885680bbcb

Observation d9107452-ab73-4e50-975e-d078c9c77215 · outbound

This paper cites Dechevski and Lars Erik Persson.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Dechevski and Lars Erik Persson

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.455466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:51.198586Z digest=sha256:78bfdf6bfc9e4f3af7d9221b7e78ba703cf26763efb758ab4d17e12f4fbdc8fa

Observation 11d9b855-377c-420a-828b-ad632b70dc71 · outbound

This paper cites On the convergence of model free learning in mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the convergence of model free learning in mean field games

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.446518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:51.317948Z digest=sha256:51a9715fb62bb98353652ba039ea0d436ff6f1e13ea0d81dbd49b17c1d490489

Observation 50967019-6c8a-4d3c-b346-641383314127 · outbound

This paper cites Learning sparse graphon mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning sparse graphon mean field games

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.438140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:51.437786Z digest=sha256:3a68c3d4a933bf16cb361e86ffd31ee67fc64ebd73bee1e38cd941a0424089fc

Observation 5b96362f-5e25-461f-9226-f9eaae412486 · outbound

This paper cites Global convergence of policy gradient methods for the linear quadratic regulator.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Global convergence of policy gradient methods for the linear quadratic regulator

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.430093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:51.590918Z digest=sha256:bf00cce61de8c29f9dcb46a0ba3973b67a053c909e4da4854b6b75fb208761c4

Observation 3f191497-aaaa-4169-b700-ae468df838aa · outbound

This paper cites Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:51.664862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:51.664862Z digest=sha256:6d7df6404cd89460bf2cd2facedaddcbd66dee1db59c1d1e817a6027b6f3a958

Observation ed498024-705a-41c7-9624-ad7968ee1dc1 · outbound

This paper cites Caines, and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines, and Minyi Huang

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.421695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:51.764041Z digest=sha256:8e04863ee1725d5876265ddf84e5c992901cb2d3f874c41a388d77767cbfb773

Observation 6e630b54-f70a-4e26-bdb3-8a924e38c07d · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:58.413846Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:51.831285Z digest=sha256:b94d0f6233ffd1802a8155617dce6ebed040d4fd0913fe78791b48fd0ccbcf8f

Observation e710deb9-6db8-4643-b121-a361319ae50d · outbound

This paper cites Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Convergence of policy gradient methods for finite-horizon exploratory linear-quadratic control problems

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.405448Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:51.925261Z digest=sha256:0b6092151e7fbf173cf38e54b06ff9b41d8cb51dfed886c1b7831d934f093ce8

Observation 4428bf96-8c76-4a93-b366-ee0bc5c299ba · outbound

This paper cites Volterra Integral and Functional Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Volterra Integral and Functional Equations

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.397411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.040532Z digest=sha256:813050c296d2c15c6cb0377cf98c498fbd9892c79673e07254e59aab8c7f2c5a

Observation d7e87767-315b-4220-ae25-0eb85ec081a4 · outbound

This paper cites Learning mean-field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning mean-field games

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.389457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.121901Z digest=sha256:9502aa5c052ba9658c5539252de9a51dcdd0ed02fc790063e5be63b884429531

Observation 7367bb92-5693-49f8-989d-86d41753bcb4 · outbound

This paper cites Policy gradient methods for the noisy lin- ear quadratic regulator over a finite horizon.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods for the noisy lin- ear quadratic regulator over a finite horizon

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.381270Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.207116Z digest=sha256:67bf5d5bfdddbf225b8a3bdaee55f4641abbb268c9fc93156c8f294e14198667

Observation 7b08fa9a-7303-454a-82f9-e396cc4d831c · outbound

This paper cites Policy gradient methods find the nash equilib- rium in n-player general-sum linear-quadratic games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods find the nash equilib- rium in n-player general-sum linear-quadratic games

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.372977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.308626Z digest=sha256:03e85dd6aec00d2568f2bbe5a49f32d7871f256b89e26af4b8b3672000fa1c86

Observation 7ab5e38b-98e3-41a0-b27d-789487e3e0b5 · outbound

This paper cites MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games MF-OML: Online Mean-Field Reinforcement Learning with Occupation Measures for Large Population Games

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:25:56.883606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.453170Z digest=sha256:8857f1f052c7cbfed6730cc1fd1a30a94cc2e8f474d29e8f8d76b506c141e7a1

Observation f9743ba9-84a0-4d46-a4d1-769d8d2caa33 · outbound

This paper cites Model-based RL for mean-field games is not statistically harder than single-agent RL.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Model-based RL for mean-field games is not statistically harder than single-agent RL

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.364277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.547369Z digest=sha256:5689979f0b34e18cdbdd7fc8d53d51dce63b68d098f16e8f585c557175c6427a

Observation f5a3cbd0-5d59-42d6-8dde-eb4278c8ae55 · outbound

This paper cites On the statistical efficiency of mean-field reinforcement learning with general function approximation.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the statistical efficiency of mean-field reinforcement learning with general function approximation

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.355754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.656118Z digest=sha256:c75745bf2b797eb05078e90fdd9fe5d4b63b461543d3126d2728c27dd945c461

Observation 95c6d6a1-48f8-4431-a192-d11babe7cdce · outbound

This paper cites Malham´ e, and Peter E.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Malham´ e, and Peter E

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.347534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.806221Z digest=sha256:aae0d5caa1f178de2502e04b5c175f10bc8e2cf541324a468fd57b4b612805a9

Observation d97d0940-69cf-4b85-b549-70b4643ccd2c · outbound

This paper cites Reinforcement Learning for SBM Graphon Games with Re-Sampling.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Reinforcement Learning for SBM Graphon Games with Re-Sampling

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T10:25:56.722759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:52.929133Z digest=sha256:4ccc2c6f524c017f6bf149229dd49c474ed14990d3c3613c9c8eac2a2ac6b0ea

Observation f49b81b1-5b32-4582-a3a9-742c8bc91cd0 · outbound

This paper cites Real-time bidding with multi-agent reinforcement learning in display advertising.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Real-time bidding with multi-agent reinforcement learning in display advertising

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.339344Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:53.031245Z digest=sha256:2a74f371a4bcb3b0362ee751b87ec4891d6ea6542918b63f127113c83e56f302

Observation 7ed1f384-d10b-4eb2-b5cf-2add752210e1 · outbound

This paper cites A natural policy gradient.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A natural policy gradient

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.331341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:53.182249Z digest=sha256:a8833347d45cb0d217e90c15d373f48677b1891ea831ad6973017b4ab12b8948

Observation 6de26ac7-ccc6-42f6-924e-b4f45d1856cd · outbound

This paper cites A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A Fisher-Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:53.289607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:53.289607Z digest=sha256:116cae75bdb497414c35443e57fa9c0dfb3a87f9ba3d872f25950f6ffb44ac87

Observation 62c2cc9f-ef5f-495d-ac4e-0dee4c654e42 · outbound

This paper cites Actor-critic algorithms.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Actor-critic algorithms

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.323352Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:53.374843Z digest=sha256:acc83f69a5a5f0971b45061a30ee29cace0b6160d2765b2a1e04b857ab1cf171

Observation 858b1adf-c7f9-4d50-8379-7af58a46ea6c · outbound

This paper cites A label-state formulation of stochastic graphon games and approximate equilibria on large networks.Mathematics of Operations Research, 48:1811–2382, 2022.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games A label-state formulation of stochastic graphon games and approximate equilibria on large networks.Mathematics of Operations Research, 48:1811–2382, 2022

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.314990Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:53.459363Z digest=sha256:f267ef55a53e50738e2494d1c8b0e3f904e82d9e7d74aa7d98920e1f6e9de15e

Observation 13e0aa9e-5bce-43cd-8a38-1ac86aa8ce23 · outbound

This paper cites Mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Mean field games

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.306776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:53.577904Z digest=sha256:ac7d8fc4c39890cebffc79b4202e10c4f94a08a59d0078335d4ce83351a88268

Observation a1055e9f-55a4-4063-b2be-b4f652e0f15e · outbound

This paper cites Learning in Mean Field Games: A Survey.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning in Mean Field Games: A Survey

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:53.713354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:53.713354Z digest=sha256:ff4187eb8cfbbe1b412ce1a1d79276169e600fc0d75abd22e6373c833b791f7e

Observation acc9d4f3-cd17-48f8-8725-88b75e370f2f · outbound

This paper cites American Mathematical Society Col- loquium Publications.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games American Mathematical Society Col- loquium Publications

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.296721Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:53.816578Z digest=sha256:d2372deef2817ef1bf5d65c4f17892fe3099df04a41aa6998c05c90b38dda6ce

Observation 1efd3b1b-a6e1-41c3-b455-ca7b01153109 · outbound

This paper cites On the global con- vergence rates of softmax policy gradient methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games On the global con- vergence rates of softmax policy gradient methods

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.288192Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:53.946606Z digest=sha256:b0f5e2e74aeed9ae73fc27e66397cd47e03da4bd57f71793ddb79927b9997b1a

Observation 21362dd4-02e6-4ad6-a1d4-04b54f59df70 · outbound

This paper cites Stochastic Graphon Games with Memory.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic Graphon Games with Memory

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.024854Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.024854Z digest=sha256:ddd9062782fc2abe7c7b7176c9ff050b4479c03ef91bc1808ab06e1af138bf17

Observation 4bb99848-22c8-4cac-90d6-99b885720120 · outbound

This paper cites Dota 2 with Large Scale Deep Reinforcement Learning.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Dota 2 with Large Scale Deep Reinforcement Learning

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.194986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.194986Z digest=sha256:d3cfaa7fd62b8c3b938f24641dd6659db30174ad95b4df9029bf37419fadb1d6

Observation c916ef8f-1cbe-44b6-98c1-b40aead4983d · outbound

This paper cites Graphon games: A statistical framework for network games and interventions.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Graphon games: A statistical framework for network games and interventions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.280044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:54.367333Z digest=sha256:a7c65b486a1c2c873481fee4744339603d16b6ab59158e9031fa40a46c88785b

Observation 06ae224f-521a-4a7a-af7f-51273bf7b88e · outbound

This paper cites Time discretization-invariant safe action re- petition for policy gradient methods.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Time discretization-invariant safe action re- petition for policy gradient methods

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.271175Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:54.597645Z digest=sha256:46f037bf7da8942c178d6c25c2a6f24b752676a3f6552e6fac5b42e260792243

Observation 8086cd2c-b1bb-41a8-b175-f2c040dfb26c · outbound

This paper cites Entropy annealing for policy mirror descent in continuous time and space.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Entropy annealing for policy mirror descent in continuous time and space

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.770111Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.770111Z digest=sha256:0f37b1886f397f6a14a50d5787f440b9ec25637623804b9e9fd923d7a1f37c67

Observation e8ae0642-33bf-4a24-be32-f7021115328a · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:54.893054Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:54.893054Z digest=sha256:964194cd5be2ffb73925aef0ddbff4e0b20b8b98c3c8c71f358834b6848dff58

Observation 12d16d47-64f4-4f3d-be60-99f0e6d88012 · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:58.262557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:54.993284Z digest=sha256:61352a29f4dbae76e810a14f09caafc66706da6497f00cc6b30e9d237e0c71ec

Observation 38e1ed21-e657-4664-abba-9b397c27886b · outbound

This paper cites Deterministic policy gradient algorithms.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Deterministic policy gradient algorithms

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.254608Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.118458Z digest=sha256:aba9511d10a6eb8090e3cc7ed855fd58bf9891d58a4d8654806144c4a5a31a55

Observation 3c8ff049-5a9f-4a34-b805-ee9926f1a9fe · outbound

This paper cites Sutton and Andrew G.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Sutton and Andrew G

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.246089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.217096Z digest=sha256:a1f8fdccb28abbb17fc32e6338f009103e4151a85c5b02983d0f3e81de1ee533

Observation 67698907-189d-49f5-9b7e-8f15530d809c · outbound

This paper cites Policy gradient methods for reinforcement learning with function approximation.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy gradient methods for reinforcement learning with function approximation

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.237861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.288745Z digest=sha256:1a41f94ab93f2fe748e5222142662fde08239e834d34eaa044958a638e34038f

Observation 83bd1352-65d0-40a7-8246-706c1de4042e · outbound

This paper cites Caines, and Minyi Huang.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Caines, and Minyi Huang

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.229525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.380324Z digest=sha256:35fa340ddacb04ecdf79aaa1b19dc28fb2b925abf74c04543833b171eb384323

Observation 72bcac30-3a04-4b6b-8cae-edf56c682601 · outbound

This paper cites Ordinary Differential Equations and Dynamical Systems, volume 140 ofGradu- ate Studies in Mathematics.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Ordinary Differential Equations and Dynamical Systems, volume 140 ofGradu- ate Studies in Mathematics

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.220709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.502383Z digest=sha256:2f195f45451a957643ced15ddfec55dffd9e3775975dcce20b143b0fa7781b83

Observation 76ebe283-1d2d-43dd-afb8-5bd04ea4373a · outbound

This paper cites Global convergence of policy gradient for linear-quadratic mean-field control/game in continuous time.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Global convergence of policy gradient for linear-quadratic mean-field control/game in continuous time

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.211525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.600933Z digest=sha256:d3958a904994e5e853022eee7ba7a59020ece66fccbc4d4f68d3cdd5161d302a

Observation 24536580-ed8b-408c-ad84-9579b75df894 · outbound

This paper cites Learning while playing in mean-field games: Convergence and optimality.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning while playing in mean-field games: Convergence and optimality

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.202300Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.699378Z digest=sha256:d6d553ee2db9d326a7fad80bb9f99e22082e048e862857b06449ce34c096fef6

Observation 20296ccf-e8b7-4bc3-9617-ba463e04787a · outbound

This paper cites Linear-Quadratic Graphon Mean Field Games with Common Noise.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Linear-Quadratic Graphon Mean Field Games with Common Noise

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T10:25:55.825786Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:25:55.825786Z digest=sha256:346fa8b3b6830e8b41d240ccea9461cdbe4b546a1f7348c7ae1ad97a7237f4db

Observation 4b5344fa-0afc-4adc-a7d2-c3c2938c86ad · outbound

This paper cites Policy mirror ascent for efficient and independent learning in mean field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Policy mirror ascent for efficient and independent learning in mean field games

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.194028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.922529Z digest=sha256:60b6de46255e52ec6b09bcd4c7afce40f18cbfdde3a27bf8dc6344ffa1982a57

Observation ff24be89-593a-455b-8716-f8dcd5f494b8 · outbound

This paper cites When is mean-field reinforcement learning tractable and relevant? In Proceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems , page 2038–2046.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games When is mean-field reinforcement learning tractable and relevant? In Proceedings of the 23rd International Conference on Autonomous Agents and Multiagent Systems , page 2038–2046

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.184842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:55.991887Z digest=sha256:4df5cd046836047dc9ea8bfcd7ab939bc72b2f9c65c3c4ca22ae02f91b549138

Observation c9ce61b0-55e5-4ec9-8000-0390774079cd · outbound

This paper cites Stochastic Controls: Hamiltonian Systems and HJB Equations.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Stochastic Controls: Hamiltonian Systems and HJB Equations

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:58.041054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:56.087561Z digest=sha256:31284c9fa5153fefd6c31aea098c5fd03cca656775c71425f3d9263b72d70d86

Observation aeac5648-6171-4005-80f5-15a06d9647db · outbound

This paper cites Learning regularized monotone graphon mean-field games.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Learning regularized monotone graphon mean-field games

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.702449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:56.207569Z digest=sha256:476ed88041b8d879d7cbf3cb5e1636d511484e442fe28db0495859ae7573ec84

Observation 63aba5c9-8a33-473d-ae08-5410e8861d43 · outbound

This paper cites an unresolved cited work.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Unresolved cited work

Reference 61

Resolution
unresolved
raw_fallback, observed 2026-08-07T10:25:57.444249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:56.268960Z digest=sha256:39f7d533e6848f0fbd45c56eee3eefbf3223ad2b7891a0090e47b532d74bb51e

Observation bef945bd-b82f-4aea-b744-4dc9ce4270aa · outbound

This paper cites Convergence of policy gradient for stochastic linear quad- ratic optimal control problems in infinite horizon.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Convergence of policy gradient for stochastic linear quad- ratic optimal control problems in infinite horizon

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.137627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:56.358470Z digest=sha256:97f27b0db84dbadafe8210fc29678731a61d91080115d18b31a291e432d16063

Observation dcf4c283-7e19-4b07-8b49-44d3999aefae · outbound

This paper cites Graphon mean field games with a representative player: Analysis and learning algorithm.

Policy Optimization for Continuous-time Linear-Quadratic Graphon Mean Field Games Graphon mean field games with a representative player: Analysis and learning algorithm

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T10:25:57.055423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T10:25:56.452058Z digest=sha256:0b28b6a6022216b34ea807a8ebb8f31351226a551d86ed2c2040f75f5d55429f

Pith citing papers

No inbound Pith citation observations are available.