Pith. sign in

Paper Citation Record · LEDGER

Policy Gradient for Continuous-Time Mean-Field Control

As of 10 August 2026, this Paper Citation Record lists 82 of 82 outbound references and 1 inbound Pith citation observation for arXiv:2605.20718.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.20718 v1

Coverage vector

measured 82 of 82 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-21T04:15:32.156380Z

measured 83 of 83 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T07:42:38.840826Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

82 of 82 outbound references displayed

  • verified exact6
  • verified fuzzy70
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 05de4c07-224a-4306-8b9a-8bf16bf4161c · outbound

This paper cites Mean field type control with congestion.Applied Mathematics & Optimization, 73(3):393–418.

Policy Gradient for Continuous-Time Mean-Field Control Mean field type control with congestion.Applied Mathematics & Optimization, 73(3):393–418

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.490060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:90cbe7a69899785b28742b8c7fcceaeb9ebcf5986f278e5f77d0384edaf181f8

Observation 3dd960af-38e9-4a6d-aaf1-1032088a9090 · outbound

This paper cites Mean field type control with congestion II: An augmented lagrangian method.

Policy Gradient for Continuous-Time Mean-Field Control Mean field type control with congestion II: An augmented lagrangian method

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.517720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:21ba062fce86127654c530dc3aed7a785c481d968c5e7536c195da5a819159c7

Observation 6fa60bf4-8ce6-44a1-a761-ee849b12035e · outbound

This paper cites A maximum principle for SDEs of mean-field type.Applied Mathematics and Optimization, 63(3):341–356.

Policy Gradient for Continuous-Time Mean-Field Control A maximum principle for SDEs of mean-field type.Applied Mathematics and Optimization, 63(3):341–356

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.499691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:5ec83c190783f296710fad8507c489c37571302e73ffb6ff872e485e86b3d552

Observation ef6e21f7-309a-4258-be77-d44777e1c7d4 · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.497337Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:006f430ee96d6527a379e4b3082cbcfd10296796415f403eaecefbc997ae5155

Observation 5d337347-a7c9-4a61-ab40-a7df3afb61e2 · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.506991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:bf311f2885fa2c9fa79dfc1f55bfe8e16fd3bc2a3899a8e6dcb51edee9c89cb8

Observation cde67c11-0658-408a-8749-49c08e5e5535 · outbound

This paper cites A weak martingale approach to linear-quadratic McKean–Vlasov stochastic control problems.Journal of Optimization Theory and Applications, 181(2):347–382.

Policy Gradient for Continuous-Time Mean-Field Control A weak martingale approach to linear-quadratic McKean–Vlasov stochastic control problems.Journal of Optimization Theory and Applications, 181(2):347–382

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.510647Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:8b2cfc67e4f9c5aca69dd292be7d1c54c02fe12994374bc6f6e7722cbe59dabe

Observation ec294a9e-112a-4809-a0db-58f70fa30209 · outbound

This paper cites Viscosity solutions of fully second-order hjb equations in the wasserstein space.SIAM Journal on Control and Optimization.

Policy Gradient for Continuous-Time Mean-Field Control Viscosity solutions of fully second-order hjb equations in the wasserstein space.SIAM Journal on Control and Optimization

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.515366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1de468128fda8019cac3e22d2b09f0fadfa7a543aacdf32dfe90d7b6b182b5f2

Observation e692bfda-9e9b-44a6-beee-57e0867ade51 · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.478246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:5d9f461da147da16b4e0cff907e6ad7365ef7e5d70657851fe68b090beb1feb9

Observation df652d8d-1a42-4a1a-813f-aaa5787aefe4 · outbound

This paper cites Comparison for semi-continuous viscosity solutions for second order pdes on the wasserstein space.Journal of Differential Equations, 455:1–31.

Policy Gradient for Continuous-Time Mean-Field Control Comparison for semi-continuous viscosity solutions for second order pdes on the wasserstein space.Journal of Differential Equations, 455:1–31

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.394800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:480efe2fe475336ee4a9ae5ce05e64daa762b6b54c3749d7dbc098a75baa4086

Observation b4866fb1-ecc5-4341-9b5a-1471cad57481 · outbound

This paper cites Comparison of viscosity solutions for a class of second-order pdes on the wasserstein space.Communications in Partial Differential Equations, 50(4):570–613.

Policy Gradient for Continuous-Time Mean-Field Control Comparison of viscosity solutions for a class of second-order pdes on the wasserstein space.Communications in Partial Differential Equations, 50(4):570–613

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.378667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:6929540041c004e151fee8b06d2310b13bfe1da83e94228f680f5bde3de109d8

Observation d168fa9b-0f8d-4a70-840a-3b6981181431 · outbound

This paper cites Convergence rate of particle system for second-order pdes on wasserstein space.SIAM Journal on Control and Optimization, 63(3):1515–1782.

Policy Gradient for Continuous-Time Mean-Field Control Convergence rate of particle system for second-order pdes on wasserstein space.SIAM Journal on Control and Optimization, 63(3):1515–1782

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.520327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:2e1aa144a644ff1495123e71dbc747d808ed492d7027352ed155ecc6ac3d1413

Observation 4dec7e8d-8736-4516-aa37-1f5df937b1ee · outbound

This paper cites Mean-field phibe: Continuous-time mean-field reinforcement learning from discrete-time data.Preprint.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field phibe: Continuous-time mean-field reinforcement learning from discrete-time data.Preprint

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.392078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:c55557d89a8590b643ebe5eafee9d95254dc0f3dd46d2a1c7bad9904f698a187

Observation bcad8097-7337-4ea7-b363-fbe4ad755c8e · outbound

This paper cites Ergodicity and turnpike properties of linear-quadratic mean field control problems.

Policy Gradient for Continuous-Time Mean-Field Control Ergodicity and turnpike properties of linear-quadratic mean field control problems

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.399715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:abb1e6e9413e690bd2956473b5548e18af3f14146cddad53bac102eb362c1d6a

Observation 225173af-5163-41b6-b458-e17f4b74a673 · outbound

This paper cites Convergence and turnpike properties of linear-quadratic mean field control problems with common noise.arXiv preprint arXiv.

Policy Gradient for Continuous-Time Mean-Field Control Convergence and turnpike properties of linear-quadratic mean field control problems with common noise.arXiv preprint arXiv

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.470720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1751af45023d38de166b6578cb12c8974c1fdc27f222648bf6bb02ab9f75212e

Observation 39ab7bfb-4943-4faf-a496-2c4a4d101bc8 · outbound

This paper cites Learning with linear function approximations in mean-field control.Journal of Machine Learning Research, 26(192):1–53.

Policy Gradient for Continuous-Time Mean-Field Control Learning with linear function approximations in mean-field control.Journal of Machine Learning Research, 26(192):1–53

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.527805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:d4389ab8babdc1546d16e022424e399d808d3b6bab71fc399025a3d2b5588e76

Observation d7f1b6b7-27fb-4d45-a2ca-6e6e7c752ad2 · outbound

This paper cites Propagation of chaos of forward-backward stochastic differential equa- tions with graphon interactions.Applied Mathematics and Optimization, 88(1):Article 25.

Policy Gradient for Continuous-Time Mean-Field Control Propagation of chaos of forward-backward stochastic differential equa- tions with graphon interactions.Applied Mathematics and Optimization, 88(1):Article 25

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.539194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:709c79ad5f9786d140038ed4279f6d79e0d9f0ff1d5257d34714ebefc02363b7

Observation d207f5f3-abf6-4604-883a-e28d5d600a40 · outbound

This paper cites SpringerBriefs in Mathematics.

Policy Gradient for Continuous-Time Mean-Field Control SpringerBriefs in Mathematics

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.541373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:5d1eab34630dbc5c0569b83f2e89a56bed9eca82fa765c6d0aabf25243407d98

Observation 8808a4bd-8df7-4f47-9962-32a1d05c70bc · outbound

This paper cites The pontryagin maximum principle in the wasserstein space.Calculus of Varia- tions and Partial Differential Equations, 58:1–36.

Policy Gradient for Continuous-Time Mean-Field Control The pontryagin maximum principle in the wasserstein space.Calculus of Varia- tions and Partial Differential Equations, 58:1–36

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.543446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:ef8a7b3056c9471d5e24f346c85e5b3dd473cff561598b1e088d269f8a5629ea

Observation 82a92baa-4dc1-41c9-a578-9f3ef272a291 · outbound

This paper cites A general maximum principle for SDEs of mean-field type.Applied Mathematics and Optimization, 64(2):197–216.

Policy Gradient for Continuous-Time Mean-Field Control A general maximum principle for SDEs of mean-field type.Applied Mathematics and Optimization, 64(2):197–216

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.545572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:9a3b90eba70046b331ce8f5ee37e05503004b1bf2e564f877dcf9ea9d65f0422

Observation ba044ed8-9990-4128-b5b3-6aef188d1762 · outbound

This paper cites Mean-field stochastic differential equations and associated pdes.Annals of Probability, 45(2):824–878.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field stochastic differential equations and associated pdes.Annals of Probability, 45(2):824–878

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.552625Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:ec2704a4969a219ebb28b8afcf54347e34b7b537ec2b3fc893e5b6609121375c

Observation 96513dbb-d7f9-4d98-9e53-619747b34459 · outbound

This paper cites Max Reppen, and H.

Policy Gradient for Continuous-Time Mean-Field Control Max Reppen, and H

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.554409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:22609f9f8dc35b32ce626d6292cea05b2c83293cdb4dfcae2aab805f0ef0e7f2

Observation 28402ac9-17e9-4ff0-a63d-9e00d9396c1c · outbound

This paper cites Society for Industrial and Applied Mathematics (SIAM), Philadelphia, PA.

Policy Gradient for Continuous-Time Mean-Field Control Society for Industrial and Applied Mathematics (SIAM), Philadelphia, PA

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.492668Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1019b42cc9c485251a264cd4153330768a634507455d374476ecd3c8bf6add99

Observation cd547b01-babd-46d0-954b-242c9ce3612d · outbound

This paper cites Forward–backward stochastic differential equations and controlled McKean– Vlasov dynamics.Annals of Probability, 43(5):2647–2700.

Policy Gradient for Continuous-Time Mean-Field Control Forward–backward stochastic differential equations and controlled McKean– Vlasov dynamics.Annals of Probability, 43(5):2647–2700

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.512969Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:3bc4a8648164f1769577757a7f4e03e467c4d4388c4f7d5f30dbb5bfb8454f9c

Observation b49508fa-a66b-437d-9356-dfcd4c8e5a95 · outbound

This paper cites I, volume 83 of Probability Theory and Stochastic Modelling.

Policy Gradient for Continuous-Time Mean-Field Control I, volume 83 of Probability Theory and Stochastic Modelling

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.564121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:64f2b10cad97cc64c87bc8da3f77f521ce6dc9a31f14501135ec2b3ca871580c

Observation 9c8ac905-bf30-40a9-8e0b-c0e701503751 · outbound

This paper cites II, volume 84 of Probability Theory and Stochastic Modelling.

Policy Gradient for Continuous-Time Mean-Field Control II, volume 84 of Probability Theory and Stochastic Modelling

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.460554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:829713352cbd2599037af506ccde402aef806e6d6251abebaf14677dc8feabc4

Observation 33e76a54-c660-4cfc-81fa-eb33ed096991 · outbound

This paper cites Control of McKean–Vlasov dynamics versus mean field games.

Policy Gradient for Continuous-Time Mean-Field Control Control of McKean–Vlasov dynamics versus mean field games

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.462811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:3eebb6461324480a66560446190be33e0e16da148279ec4f9c438267a19e1ee4

Observation 3ec279bb-67a8-4c4e-bb19-ac77d2674b18 · outbound

This paper cites Mean field games and systemic risk.Communications in Mathematical Sciences, 13(4):911–933.

Policy Gradient for Continuous-Time Mean-Field Control Mean field games and systemic risk.Communications in Mathematical Sciences, 13(4):911–933

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.467872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f036d5d6ec024c5ff567ffc2969eb1dc9c060aa0c5281c7bfbe1aa6159fe4c63

Observation 5db4fbbe-c53a-4e30-916c-e4fb5ee5636b · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.536369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:ec889741fa961433cc382895259283eaae8eb8e3241c2705fc7465abe82fb0a3

Observation 48c6c7de-978a-48f0-b6ba-e4f829410715 · outbound

This paper cites Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods.

Policy Gradient for Continuous-Time Mean-Field Control Linear-Quadratic Mean-Field Reinforcement Learning: Convergence of Policy Gradient Methods

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.507409Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:2681945849a0276c5333a792162958d9ee2c146919f931cd22db7a46a7ba59cc

Observation d99aaeb9-1669-40bc-9057-5e59ec113a97 · outbound

This paper cites Carrillo, Young-Pil Choi, and Maxime Hauray.

Policy Gradient for Continuous-Time Mean-Field Control Carrillo, Young-Pil Choi, and Maxime Hauray

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.397202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:65e03f043ae5a748027b396050475246e61a177df770a8dc8c9743a3749ad6b9

Observation 0e36a81f-ea3d-4d28-808f-6a37742f26e7 · outbound

This paper cites Carrillo, Massimo Fornasier, Giuseppe Toscani, and Francesco Vecil.

Policy Gradient for Continuous-Time Mean-Field Control Carrillo, Massimo Fornasier, Giuseppe Toscani, and Francesco Vecil

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.457906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:9f6fb04b514f761f93065847915fee068d7b4dd5cb901146b8126b600c46d29f

Observation c4f9cfa8-aaae-4a02-b634-f9e9ab96c9d1 · outbound

This paper cites Propagation of chaos: A review of models, methods and applications.

Policy Gradient for Continuous-Time Mean-Field Control Propagation of chaos: A review of models, methods and applications

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.559847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:a0e81a42f4440770b0c5c0106922cd1241dd2aeb3539b2da67529be6841e6df5

Observation 4f1a21e8-3e36-4c01-9bb5-76742afaa769 · outbound

This paper cites Numerical method for FBSDEs of McKean–Vlasov type.Annals of Applied Probability, 29(3):1640–1684.

Policy Gradient for Continuous-Time Mean-Field Control Numerical method for FBSDEs of McKean–Vlasov type.Annals of Applied Probability, 29(3):1640–1684

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.473218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:bbd38ea164f135e6d5d0159b7b285558604f007db88ff0a5e0d7299ecca57b3c

Observation 9cdb784c-2560-41da-8ed4-3fca87cd6654 · outbound

This paper cites Emergent behavior in flocks.IEEE Transactions on Automatic Control, 52(5):852–862.

Policy Gradient for Continuous-Time Mean-Field Control Emergent behavior in flocks.IEEE Transactions on Automatic Control, 52(5):852–862

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.447406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:b51e1c28eae763343a14ada0b64ee256c61bc6a2df2b14c48e219dfe5db2b100

Observation 53120db2-2bdb-42f9-935b-c6334b7ed650 · outbound

This paper cites Mckean–vlasov optimal control: Limit theory and equivalence between different formulations.Mathematics of Operations Research, 47(4):2891–2930.

Policy Gradient for Continuous-Time Mean-Field Control Mckean–vlasov optimal control: Limit theory and equivalence between different formulations.Mathematics of Operations Research, 47(4):2891–2930

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.452538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:d45fe91b24d93dabc4656ca697730de130a56ef71c013bbbd60a690309188140

Observation e6713b11-f19b-481a-a0dc-f6ed3517fff9 · outbound

This paper cites Martingale measures and stochastic calculus.Probability Theory and Related Fields, 84(1–2):83–101.

Policy Gradient for Continuous-Time Mean-Field Control Martingale measures and stochastic calculus.Probability Theory and Related Fields, 84(1–2):83–101

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.444803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:7bbe8b110fce21fcf576b48c6bb1324fd33a481b28276b2c9a3715444eaefa16

Observation e0e10495-ee5a-4664-8447-ddaa699b2b09 · outbound

This paper cites Actor-critic learning for mean-field control in continuous time.J.

Policy Gradient for Continuous-Time Mean-Field Control Actor-critic learning for mean-field control in continuous time.J

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.449933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:da8df3eab06845dc41acf0a3aaf8636ccf2b7501e3111fb9b3c0c75a1e6df146

Observation b91cba99-3aef-4121-adbf-7a019cae5bc3 · outbound

This paper cites Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise.

Policy Gradient for Continuous-Time Mean-Field Control Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.522026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:b9ca4b9cc30e5247dd5842bf325f2439409603c31217ab7a10f8e4765253ca91

Observation 4ab26cd6-e292-4b09-871d-98e5010e1d6e · outbound

This paper cites Hamilton-Jacobi equations in the Wasserstein space.Methods Appl.

Policy Gradient for Continuous-Time Mean-Field Control Hamilton-Jacobi equations in the Wasserstein space.Methods Appl

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.561911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:e0953b587cc5db7ba92b67e2c4492fd8104e6ce83807f07478616627d85749bc

Observation 59dcfc00-79c1-44c8-a0ed-aab8e9abf1d5 · outbound

This paper cites Opinion dynamics and bounded confidence: Models, analysis and simulation.

Policy Gradient for Continuous-Time Mean-Field Control Opinion dynamics and bounded confidence: Models, analysis and simulation

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.436610Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:55407a68885eb65bd640474d1284281e165b63ce93e488428e1a44d811992af1

Observation 498872b3-5e19-41c8-a718-7b7eff381ecf · outbound

This paper cites Howard.Dynamic programming and Markov processes.

Policy Gradient for Continuous-Time Mean-Field Control Howard.Dynamic programming and Markov processes

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.404624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:7f2df10aa808248721927e826da2d56e69c4c92126227f60bea5419dfac59799

Observation b2d33492-3963-43e9-b498-24bc11a715c3 · outbound

This paper cites A linear-quadratic optimal control problem for mean-field stochastic differential equations in infinite horizon.Math.

Policy Gradient for Continuous-Time Mean-Field Control A linear-quadratic optimal control problem for mean-field stochastic differential equations in infinite horizon.Math

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.475956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1338d1d3d8beba8ecbd5421eec35ceab75c55a96c10ec1250a7d5c4a118bb2d6

Observation a31369c5-5175-4a8b-a28a-a19960545c0d · outbound

This paper cites Infinite horizon value functions in the Wasserstein spaces.J.

Policy Gradient for Continuous-Time Mean-Field Control Infinite horizon value functions in the Wasserstein spaces.J

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.434217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:94de5ce4bbcc9c908159cbc741ff92cd0b2e21c67678f00429396648d11dadb5

Observation d9d8dd3c-e052-4017-aee4-6d552d7b7668 · outbound

This paper cites Accuracy of discretely sampled stochastic policies in continuous-time reinforcement learning.

Policy Gradient for Continuous-Time Mean-Field Control Accuracy of discretely sampled stochastic policies in continuous-time reinforcement learning

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.511413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:79950179b900ca9e16f156c80d3fbd80bfe8b5c8db6233103d429f6b321b588e

Observation 99f04930-10d8-40bc-bc74-82556801903d · outbound

This paper cites Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach.Journal of Machine Learning Research, 23(1):6918–6972.

Policy Gradient for Continuous-Time Mean-Field Control Policy evaluation and temporal-difference learning in continuous time and space: A martingale approach.Journal of Machine Learning Research, 23(1):6918–6972

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.522899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f2e989f60614afaafb06129fec975dd96ed6c81f3700c69acc298f33f8ce0c8f

Observation 1c89e2b0-8c49-498e-b000-4347767a5658 · outbound

This paper cites Policy gradient and actor-critic learning in continuous time and space.Journal of Machine Learning Research, 23(1):12603–12652.

Policy Gradient for Continuous-Time Mean-Field Control Policy gradient and actor-critic learning in continuous time and space.Journal of Machine Learning Research, 23(1):12603–12652

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.480561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:49cb7b60e9e47a074bc299772fc23e52f134beb9ae15c4a350f91059c1d62f78

Observation be968192-42a4-443c-be7d-a8692dc2f708 · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.428864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:cf38a558ee015ec9bf7b8d7fd4624da2bca83e9b9abe1ce07118fcdf0ac82da5

Observation 08a394e0-e2db-4bbf-b050-766fcd1d206e · outbound

This paper cites an unresolved cited work.

Policy Gradient for Continuous-Time Mean-Field Control Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-21T10:04:59.529917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:d0ce43ae932c5bcea3eaf3705cb36b2c69a86b9a44b4293a34cd6b75be7bfbd4

Observation 17085b17-2973-44a8-b454-a1f6ccb74323 · outbound

This paper cites Limit theory for controlled McKean–Vlasov dynamics.SIAM Journal on Control and Optimization, 55(3):1641–1672.

Policy Gradient for Continuous-Time Mean-Field Control Limit theory for controlled McKean–Vlasov dynamics.SIAM Journal on Control and Optimization, 55(3):1641–1672

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.487899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:3ca4d0161a28d27df30b835cf6e8fc57e375de3c66a271b65f4b966a0ef9a4fe

Observation b1a2b330-50da-41ab-8bb1-ddb4cc4f77d2 · outbound

This paper cites Mean field games.Japanese Journal of Mathematics, 2(1):229–260.

Policy Gradient for Continuous-Time Mean-Field Control Mean field games.Japanese Journal of Mathematics, 2(1):229–260

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.482972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:85bf3a47418660a1f6480c73699965802927d02a7361e524f28987ef4c933fe6

Observation 6871f50e-e74e-466b-bab9-4bbf6a2e86d8 · outbound

This paper cites Numerical methods for mean field games and mean field type control.Proceedings of Symposia in Applied Mathematics, 78:221–282.

Policy Gradient for Continuous-Time Mean-Field Control Numerical methods for mean field games and mean field type control.Proceedings of Symposia in Applied Mathematics, 78:221–282

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.534428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:847d19c0f4fcec18bc99df20de20d09d685ee8c27418d506f78aeb526b917b6a

Observation 8b2ef5be-d50f-4397-9b32-d6569ec2a66c · outbound

This paper cites Dynamic programming for mean-field type control.Comptes Rendus Math´ ematique.

Policy Gradient for Continuous-Time Mean-Field Control Dynamic programming for mean-field type control.Comptes Rendus Math´ ematique

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.532000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:e462870090afc19570e812234c7c157b7500d5ca74f75459bc0e24665c677cbb

Observation 750f29c2-a508-46ea-a5f9-d833186900ac · outbound

This paper cites Mean-field stochastic linear quadratic optimal control problems: Closed-loop solvability.Probability, Uncertainty and Quantitative Risk, 1(1):2.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field stochastic linear quadratic optimal control problems: Closed-loop solvability.Probability, Uncertainty and Quantitative Risk, 1(1):2

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.485361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:13d3b21cc345413553809bf2616f364454c95b0b3dc3b6fddf9df5b7ca171aae

Observation a896a42a-f91f-4b0f-974c-5f53b29afb26 · outbound

This paper cites Cours au Coll` ege de France: Th´ eorie des jeux ` a champ moyen.

Policy Gradient for Continuous-Time Mean-Field Control Cours au Coll` ege de France: Th´ eorie des jeux ` a champ moyen

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.426546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:5451f6aa9a56303bc29feb3c97ac23333ea6b917391f3b3ca363e5b8b0940f7b

Observation c6f34b89-825b-4fcc-8308-0c01065439f3 · outbound

This paper cites Linear quadratic optimal control of conditional McKean–Vlasov equation with random coefficients and applications.Journal of Mathematical Economics, 66:7–26.

Policy Gradient for Continuous-Time Mean-Field Control Linear quadratic optimal control of conditional McKean–Vlasov equation with random coefficients and applications.Journal of Mathematical Economics, 66:7–26

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.431615Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f4bd8b8c8879fc40e038cb73d7e94a804e77d83c77819222bf270fbca3f42942

Observation a0d208c2-9953-4632-96ef-429bee3168c7 · outbound

This paper cites Actor-critic learning algorithms for mean-field control with moment neural networks.

Policy Gradient for Continuous-Time Mean-Field Control Actor-critic learning algorithms for mean-field control with moment neural networks

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.439243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:02375fdac009777442083d5c3eb36148f5d30e07862a91c8537b3f4e22753ce4

Observation 73951ded-2a0f-448b-a4c4-607421a6fb57 · outbound

This paper cites Dynamic programming for optimal control of stochastic McKean–Vlasov dynamics.

Policy Gradient for Continuous-Time Mean-Field Control Dynamic programming for optimal control of stochastic McKean–Vlasov dynamics

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.495075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1a0fd9a049cab9a35c8957d4495bbf7806bb14143e6e57a263b6823fe2532009

Observation 33d2da7e-64cc-45b0-996d-4e149cbfe98a · outbound

This paper cites Bellman equation and viscosity solutions for mean-field stochastic control problem.

Policy Gradient for Continuous-Time Mean-Field Control Bellman equation and viscosity solutions for mean-field stochastic control problem

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.419115Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:64a110c0a4158890621b700504d2dfe570ff7dba3b25a882b3d4917ca5a1b439

Observation ee922660-e942-4f26-9215-b07d7eaf58b9 · outbound

This paper cites Mean-field neural networks: Learning mappings on wasserstein space.Neural Net- works, 168:380–393.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field neural networks: Learning mappings on wasserstein space.Neural Net- works, 168:380–393

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.424023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:b0dfa7f83930e8d5a5a90ae663eb868b9196fbb96624ee0602963abf7d58f8f7

Observation 5d6cec1d-165e-412e-b94c-f66e66f1fec7 · outbound

This paper cites Continuous-time q-learning for mean-field control with common noise, part-i: Theoretical foundations.

Policy Gradient for Continuous-Time Mean-Field Control Continuous-time q-learning for mean-field control with common noise, part-i: Theoretical foundations

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.416581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:7c1a66ac30bf8fff75f676d0ea01074943c4b82bbe48050045b60bc19ea73937

Observation 5c65b131-2f3d-47b2-a3ec-500ec902646c · outbound

This paper cites Continuous-time q-learning for mean-field control with common noise, part-ii: q-learning algorithms.

Policy Gradient for Continuous-Time Mean-Field Control Continuous-time q-learning for mean-field control with common noise, part-ii: q-learning algorithms

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.421629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:0cb35131d7d5d370b0a0c51a9634743d9b6b3208868f77f6b4f0a09354fe7e57

Observation 7f8605af-8b72-47ce-80d7-d89af6336aec · outbound

This paper cites Osher, Wuchen Li, Levon Nurbekyan, and Samy Wu Fung.

Policy Gradient for Continuous-Time Mean-Field Control Osher, Wuchen Li, Levon Nurbekyan, and Samy Wu Fung

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.441828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:d50fa34ef399593c393d258fa63c2ccc622894a712d63f306e0fb2542d9dfce1

Observation 288cf3e9-8ef0-43b6-bcb8-7d44fc1e9bb3 · outbound

This paper cites Learning algorithms for mean field optimal control.

Policy Gradient for Continuous-Time Mean-Field Control Learning algorithms for mean field optimal control

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.514765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:3a6faafb3b9815ee914a1b4b7ebfd3d0de03b740ba0c28fe0ceb04781498f5f2

Observation 695a9fc3-c832-493b-bb48-f36ca5edef3e · outbound

This paper cites Mete Soner and Qinxin Yan.

Policy Gradient for Continuous-Time Mean-Field Control Mete Soner and Qinxin Yan

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.550478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:36955a17be2a030445244778b65630dd6926541a105b219f01c7b0070cd55460

Observation a27f048f-3d6b-469f-95d9-e39eb45c394e · outbound

This paper cites Mete Soner and Qinxin Yan.

Policy Gradient for Continuous-Time Mean-Field Control Mete Soner and Qinxin Yan

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.409253Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:2478cba09497a9b2a652b9bf387081ad051d88babe1dec0fd4eb0af984f9b5be

Observation 215e1ce4-f58b-40b7-81f5-bba0820c2935 · outbound

This paper cites Mean-field stochastic linear quadratic optimal control problems: Open-loop solvabilities.ESAIM: Control, Optimisation and Calculus of Variations, 23(3):1099–1127.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field stochastic linear quadratic optimal control problems: Open-loop solvabilities.ESAIM: Control, Optimisation and Calculus of Variations, 23(3):1099–1127

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.384181Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:468909474e22b5c200f0f88b473cdafc294d0103213108175e5c26a0feee6d6f

Observation 8eeaf5bf-84c3-4b26-96da-ee7a43ae9c64 · outbound

This paper cites Mean-field stochastic linear-quadratic optimal control problems: Weak closed-loop solvability.Mathematical Control and Related Fields, 11(1):47–71.

Policy Gradient for Continuous-Time Mean-Field Control Mean-field stochastic linear-quadratic optimal control problems: Weak closed-loop solvability.Mathematical Control and Related Fields, 11(1):47–71

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.381332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:ca811db182de39121a5a011a446e6720d318332336474d1f70888ab122ff3767

Observation 97ec6ec2-e7f9-4206-81c6-2dd49cdf60c4 · outbound

This paper cites Springer Nature.

Policy Gradient for Continuous-Time Mean-Field Control Springer Nature

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.455576Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:1609488a02cafa7ae4a0b504a22e2c82049d791a9d4a312a0e5f45ed9588be75

Observation 695e3a34-7ebb-4324-b980-ef4580092137 · outbound

This paper cites The exact law of large numbers via fubini extension and characterization of insurable risks.Journal of Economic Theory, 126(1):31–69.

Policy Gradient for Continuous-Time Mean-Field Control The exact law of large numbers via fubini extension and characterization of insurable risks.Journal of Economic Theory, 126(1):31–69

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.504533Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:d8e7c3a8d75288839f633db32ed891a1eaa527ba489fd77be09e58ba61f9ea33

Observation 27c5ee06-cd2e-46ce-b259-15c9c6894deb · outbound

This paper cites Sutton and Andrew G.

Policy Gradient for Continuous-Time Mean-Field Control Sutton and Andrew G

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.414273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:0dcf7c7cf311e555967684371fe9155383766dd72239dcc73824ca39fa408aa6

Observation ee42fb72-20b3-4147-ae61-4e6c5795b0c1 · outbound

This paper cites Synthesis Lectures on Artificial Intelligence and Machine Learning.

Policy Gradient for Continuous-Time Mean-Field Control Synthesis Lectures on Artificial Intelligence and Machine Learning

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.502141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:180412326928d6289ca9dad0ec8c07759e77f7b7ebf1b7bcde776202a8f602e9

Observation cb9dc937-8511-442d-901c-1fa2ce1b82d5 · outbound

This paper cites Topics in propagation of chaos.

Policy Gradient for Continuous-Time Mean-Field Control Topics in propagation of chaos

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.525348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f79ba58224866cf47f531eaa9e036134808e3c4671364227a643ccdad1e21a92

Observation db9aa6b9-6bc6-4496-bcfd-b45ad55f4b41 · outbound

This paper cites Optimal scheduling of entropy regularizer for continuous- time linear-quadratic reinforcement learning.SIAM Journal on Control and Optimization, 62(1):135–166.

Policy Gradient for Continuous-Time Mean-Field Control Optimal scheduling of entropy regularizer for continuous- time linear-quadratic reinforcement learning.SIAM Journal on Control and Optimization, 62(1):135–166

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.402198Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:596d5f6f5fa8422121f4c8654316d5c16ee600efd622e89aa639ebe16ef6fb31

Observation b5b60a16-c453-4f04-9944-54b1f1d33506 · outbound

This paper cites Making deep Q-learning methods robust to time discretization.

Policy Gradient for Continuous-Time Mean-Field Control Making deep Q-learning methods robust to time discretization

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.407104Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:f14eb02c3f83be97c5b4d5897076d106899d9b1ca9e8edde3f79c1d693f53376

Observation e7fbaa1d-bd78-42ed-af19-367c06530e3c · outbound

This paper cites Kinetic models of opinion formation.Communications in Mathematical Sciences, 4(3):481–496.

Policy Gradient for Continuous-Time Mean-Field Control Kinetic models of opinion formation.Communications in Mathematical Sciences, 4(3):481–496

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.389495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:bcc00c64cffd7fe85b53ac181fabdabbf8e0ead55615892826869fbd1f6bdaaa

Observation f5f4cf72-5e45-4958-adc9-34d24c3fe061 · outbound

This paper cites Reinforcement learning in continuous time and space: A stochastic control approach.Journal of Machine Learning Research, 21(198):1–34.

Policy Gradient for Continuous-Time Mean-Field Control Reinforcement learning in continuous time and space: A stochastic control approach.Journal of Machine Learning Research, 21(198):1–34

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.386913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:6af350e2ae20b5e6da460420aa8d9ce443117227ae8da3acdc08b7dbb3261b14

Observation b1047761-c87e-4955-a12a-fac5d9b565e0 · outbound

This paper cites Global convergence of policy gradient for linear- quadratic mean-field control/game in continuous time.

Policy Gradient for Continuous-Time Mean-Field Control Global convergence of policy gradient for linear- quadratic mean-field control/game in continuous time

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.411820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:8bd74fc15636969c3f40ae2bb4de99c5721c577d8f06d50ff9542234debae433

Observation 0093bcd8-b17e-4bc6-a2cb-298d91ca5470 · outbound

This paper cites Continuous time q-learning for mean-field control problems.Applied Mathematics & Opti- mization, 91(1):10.

Policy Gradient for Continuous-Time Mean-Field Control Continuous time q-learning for mean-field control problems.Applied Mathematics & Opti- mization, 91(1):10

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.547797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:4ec40e0dc03884185f3a973c40a9777da151925a654c1e0959c40b7224acb9f9

Observation e5db062e-971d-4d5a-86c4-ebd664ea8d69 · outbound

This paper cites Unified continuous-time q-learning for mean-field game and mean-field control problems.

Policy Gradient for Continuous-Time Mean-Field Control Unified continuous-time q-learning for mean-field game and mean-field control problems

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-08-03T01:10:22.407760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:c39c29b257ea47e19b575fcc293266c3fefdabe34b69f7316e18b55627e188b0

Observation 77da7314-7111-4e8d-a5b9-4924946a5b33 · outbound

This paper cites Efficient local planning with linear function approximation.

Policy Gradient for Continuous-Time Mean-Field Control Efficient local planning with linear function approximation

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.465245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:e237842a22b5ff60fb8c4ab0872424e3e72af13a4fb3b26c765119d56850d165

Observation 50d101de-2164-4cc4-b593-91e8ba4da176 · outbound

This paper cites Linear-quadratic optimal control problems for mean-field stochastic differential equations.SIAM Journal on Control and Optimization, 51(4):2809–2838.

Policy Gradient for Continuous-Time Mean-Field Control Linear-quadratic optimal control problems for mean-field stochastic differential equations.SIAM Journal on Control and Optimization, 51(4):2809–2838

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-05-21T10:04:59.557643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:b59edb04f60f8ac552f5512d241d624bbafc857e67a56223efd50262075a8c9a

Observation 5c5ca38c-7cdb-48f8-adcb-87334f300f85 · outbound

This paper cites PhiBE: A PDE-based Bellman equation for continuous time policy evaluation.

Policy Gradient for Continuous-Time Mean-Field Control PhiBE: A PDE-based Bellman equation for continuous time policy evaluation

Reference 82

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:19:33.525856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-21T04:15:32.156380Z digest=sha256:dca1554d27badd89e501cc22c99c15b614fab7a68794957530b39495f3a89754

Pith citing papers

Observation 4ad6043e-c0ec-4375-9534-e26bdd39c1c0 · inbound

Actor-Critic Learning for Extended Mean Field Control with Deterministic Policies cites this paper.

Actor-Critic Learning for Extended Mean Field Control with Deterministic Policies Policy Gradient for Continuous-Time Mean-Field Control

Reference 3

Resolution
unresolved
no resolver link, observed 2026-07-14T07:42:38.840826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:42:38.840826Z digest=sha256:fdd250dd753109372b741d1388a6644099e64eb24b4a5c45cea1bb928fc6a00f