Pith. sign in

Paper Citation Record · LEDGER

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms

As of 4 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 2 inbound Pith citation observations for arXiv:2604.27378.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2604.27378 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-07T08:41:31.247249Z

measured 64 of 64 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-04T06:34:03.388597+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T07:42:38.840826Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T05:29:35.391707Z

Reference resolution

62 of 62 outbound references displayed

  • verified exact13
  • verified fuzzy34
  • unresolved15
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 6fd4fa7a-784b-4257-8ebb-bf37f32b757e · outbound

This paper cites Ambrosio, N.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Ambrosio, N

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.507078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:1b6629c28d5ba9c64bdebd27494bbdbe1aaecdcd8d3c14c2ca01448b46fdfb0a

Observation 18fba9be-6349-466b-b7f0-af0ea18010f0 · outbound

This paper cites Q-Learning in Regularized Mean-field Games.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Q-Learning in Regularized Mean-field Games

Reference 2

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.915770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:4a402179af5fc47042ced10d4e52e5d38d7af212329ea48a5cc5ad9da22555f0

Observation 00918088-82e0-4570-a40d-64a0bd3f4db5 · outbound

This paper cites Angiuli, J.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Angiuli, J

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.504298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:ffaf9dd7ae0021cd9be78cf74f81b1533b965f0467bdd863ded9dd789a3ef6b7

Observation 617b0a4c-1fa5-4357-80e5-05772dfbb70c · outbound

This paper cites Deep Reinforcement Learning for Infinite Horizon Mean Field Problems in Continuous Spaces.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Deep Reinforcement Learning for Infinite Horizon Mean Field Problems in Continuous Spaces

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.899151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:c98d965a3e490dcea4d3891f5b76eb8d7bb4e671f347ddf1e377d51ca27f14b4

Observation 77caa214-d9d2-4d7b-b3fb-2db2e206c2ba · outbound

This paper cites Convergence of Multi-Scale Reinforcement Q-Learning Algorithms for Mean Field Game and Control Problems.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Convergence of Multi-Scale Reinforcement Q-Learning Algorithms for Mean Field Game and Control Problems

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.975267Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:1b7a8874795c2841d88599ba3ae027f35d176300515c9eb02845525c8a21f74f

Observation 6e259bca-d5dc-44c0-864e-fbdd92b01e52 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.487152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:1a9d376f59ee1778b2ced2a4cc0383062d88562c29a3ff5dbe61f3830604df86

Observation a5b619ba-80eb-453c-9f20-5391c8f0795a · outbound

This paper cites Carmona and F.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Carmona and F

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.490086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:744f3d8e6b91c253ed144c6aeaec24001b54b65cb742fea07a56536b97c6a7af

Observation 455dfa3f-3b0a-46ac-8f5a-14ee15426f19 · outbound

This paper cites Carmona and F.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Carmona and F

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.493502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:0af670b0b05718358cd8f049fbc1e6bc526f77a6b1807a37ee5ba781c783b4a0

Observation c6ca25dd-ddd0-4f0a-a3ba-eb82ddf9abcc · outbound

This paper cites Carmona, J.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Carmona, J

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.496589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:66d2b3b093014bd833aa43df53be0dec30554cc1d61099ca67188ba999a9f425

Observation 194604b8-f7df-43f3-a12d-85dafe5262f2 · outbound

This paper cites Carmona, F.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Carmona, F

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.469652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:cffd9035cada075e6c4e4acd4f0d1f9519d4fe836fe1911175a094478d6bc14c

Observation ce89d044-9b9b-4ce4-8b44-0c831301a5f6 · outbound

This paper cites Carmona and M.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Carmona and M

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.955564Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:f1adf60e2e744acbba0d217535b6635e6db2a3719512b31b0a8fe6bef1c17a8d

Observation 18991aed-88c4-4f70-8ac1-494dab7f6a4b · outbound

This paper cites Carmona, M.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Carmona, M

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.472408Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:a0076b1738df33709b53f6073e84653f705aa5ed4e3f7feb9180bc41d2dc372a

Observation 3adeda22-202b-4a7d-a997-5c963916840a · outbound

This paper cites Chassagneux, D.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Chassagneux, D

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.467157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:c6a5d4fccc9c9162a61804008e6b89016966a6541f98161e50289fc01833cb4f

Observation d12ddc27-c43e-4aae-8fa3-c129830563e9 · outbound

This paper cites A Viscosity Solution Theory of Stochastic Hamilton-Jacobi-Bellman equations in the Wasserstein Space.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms A Viscosity Solution Theory of Stochastic Hamilton-Jacobi-Bellman equations in the Wasserstein Space

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.904460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:762cf859e6a059e8ba82d81db641704871732e8fee23ba3eacc7ddb1ce25439e

Observation 74a35787-ce47-4a2f-ad3e-b4e9109e3fdb · outbound

This paper cites Conforti, A.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Conforti, A

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.501651Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:854560c8b041cbc6b12973ce2df75fe8fe4a3e8b1aa3ad9bfe2b31687c364ed0

Observation 6241c150-153d-45c7-a438-afa305d35a0b · outbound

This paper cites Optimal control of path-dependent McKean-Vlasov SDEs in infinite dimension.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Optimal control of path-dependent McKean-Vlasov SDEs in infinite dimension

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.910582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:30ea7a8a5bf2f7a0386990049162f22b232c4905de56cedf64ed7416f0ab59d8

Observation 653902d8-b95e-450e-b05f-fd1599e67975 · outbound

This paper cites Crisan and E.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Crisan and E

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.461348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:ac141ebde50bcdfeb45ccaa70a95a6aaf6d9fc4a9463a49a5e2bddb6b1159807

Observation f631f891-a8fc-41cd-bfe2-0adc2bd2392b · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.464398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:73a1d800d85f34785b152af42ab907c3763a29f6ed093059a6fbe1e42e0fee88

Observation e9221019-9f33-4137-b55b-8c676915f71e · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.475220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:6116ded3dced7e95c079963d4f5b6327c0e045aed61abc1a7d0c4f7b58503e07

Observation 210d4583-2cd0-448f-9d4f-151c6a4d6561 · outbound

This paper cites arXiv preprint arXiv:2312.11797 , year=.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms arXiv preprint arXiv:2312.11797 , year=

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.939953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:eb879511d2224c9cff8817552c73cd23c1c1b115dac415b90a6852f98355a1c7

Observation f31c65c6-c46c-489a-9f29-314d671e1edd · outbound

This paper cites Djete, D.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Djete, D

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.484256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:f41bebac148f32bfe0cda5f7349e6741937a7c09ccc499abbeb88531ba105041

Observation 3334f914-b3ee-47e5-a83b-233227ef7603 · outbound

This paper cites Dong (2024): Randomized optimal stopping problem in continuous time and reinforcement learning algorithm.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Dong (2024): Randomized optimal stopping problem in continuous time and reinforcement learning algorithm

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.453449Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:f86e4111c46ae16f45188d0086ebbeb4943f74eb5d2dd71c1064d748f81da8e6

Observation 5519ca4c-fc4b-4c66-b87f-5eeca044af28 · outbound

This paper cites Djete, D.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Djete, D

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.458847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:205722e38bae0d59ff0d6c63a228579de2f05f998ed79e7766f1208a2577b5e1

Observation 85500359-27cf-485e-af8a-4fcf93372542 · outbound

This paper cites Dupuis, R.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Dupuis, R

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.450330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:664b433dc6013c1a777f33a1ec6b132ad10dbca316792275b97a12236476b19a

Observation 0727156c-d34b-4fa3-ad1a-1da0dd6b9b50 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.441238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:8e025fdb99282345e181e8bbc7109f2c425d744abca6241ffccf0b34e8284640

Observation fb7b2a1f-d804-43fa-a97f-04de2e1064c7 · outbound

This paper cites Doya (2020).

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Doya (2020)

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.515203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:caceff0b5f002d7d70c574d57074b5fdbf9accb6b16f6024a76971b8748c8c7e

Observation 33352ebd-dc77-4fcf-906b-c72141662ed4 · outbound

This paper cites Firoozi and S.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Firoozi and S

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.444105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:1d9d3fc2ffb9f19f3589237c01b93f3e202ecd099d053734c29145fbe79e6904

Observation 2f3334cc-cb3a-4e03-a264-0ee885bc276f · outbound

This paper cites Frikha, M.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Frikha, M

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.447245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:199ba37e6506172ec73d9790eca9440e00d311b1ce2168142db616c6cbe5ab47

Observation 0b25da7c-95d8-442d-922b-023b715f1607 · outbound

This paper cites Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Full error analysis of policy gradient learning algorithms for exploratory linear quadratic mean-field control problem in continuous time with common noise

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.967119Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:08ecc677f2722cf647d16c88862df9114af80ddfd8b8b07d9176a6ca4a106f70

Observation f23a04ce-4c39-42de-a4ec-982c9aaa991f · outbound

This paper cites Giegrich, C.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Giegrich, C

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.456206Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:2c043b296568659e31ae9be8f50902c8d64e17f9b3327ae5b46c645ad345a5ae

Observation 5afb29cf-748f-4cb7-a3ea-59c5baad18c2 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.478226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:cf4e0d8254a396f1c3e65419a4b8a6e0b85e59d5d9a92ec6776945bdfb1ffbdc

Observation d3fbe909-a6de-4606-861e-3f9d467181cb · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.435186Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:565100a7a69c32a09ddac3a8fa1782b6605919dd73472127c91c3330941bb45a

Observation e3e3a0a5-d8ee-4cf3-86ec-39a8ac3be30a · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.432427Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:ca45584e62044f84f0b8798c0c5e02e4b6557de8a034e1b29ac1707dffaff08e

Observation 3ea8048a-60bb-4366-ab7c-8f92ac1a4d13 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.438401Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:d25f744d1f9b5903a3eb8b98f77d0a07ccd8566f1c7f8d47a5a222a5fb34c25d

Observation 4117b598-bf81-4f0b-96b1-8232ddb200c1 · outbound

This paper cites Continuous-time reinforcement learning for optimal switching over multiple regimes.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Continuous-time reinforcement learning for optimal switching over multiple regimes

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-31T02:03:07.292107Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:a65996e6e307e556e98946cde3830873f21d26a3faea99808edd0d2fdf49c40e

Observation 135c9a6a-23dc-4107-afc8-08dd8c27c00b · outbound

This paper cites Jia and X.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Jia and X

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.481369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:46f57c03e9a3b9f3026ee8b5e6413112c3263811522e2e1cc221254d69f313c1

Observation de6f15cc-5e87-4a45-ac1c-d49dc978e12f · outbound

This paper cites Jia and X.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Jia and X

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.499137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:7612fc8ddbadb2410b091de0dbdf8c2bdbd9084003e6a179f826ad5388a13a60

Observation 7b5fb4e3-ebfa-4987-9d1b-86f8e1e600f5 · outbound

This paper cites Jia and X.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Jia and X

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.571962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:e40b57a6eff605da20cc2ea156ead5bb8557398f44b75b1bb416e7d0202e5311

Observation 75aae826-5d3e-4a12-8383-7826c29a37cd · outbound

This paper cites Jia (2026): Continuous-time risk-sensitive reinforcement learning via quadratic variation penalty.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Jia (2026): Continuous-time risk-sensitive reinforcement learning via quadratic variation penalty

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.566482Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:5e9e6b7cd1fe52177b246ff8aea070e0808c0231a76feaa0d5d30849c5a7553a

Observation 9cb2366f-602e-4023-935e-357e470f9234 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.558200Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:29bc767dbba62396c50c7f43273e5506965fc017c060800552be02936de7eaa1

Observation 7998e14e-5365-42dd-8453-38800bbc0afd · outbound

This paper cites Kallenberg(2002): Foundations of Modern Probability.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Kallenberg(2002): Foundations of Modern Probability

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.560965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:32bbabc837d73a2e1523ce8f81115d0d5180aaf71acddb7e2cc3ec55df11cdec

Observation b0ebe943-095a-4702-974e-ec183573a7f1 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.563712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:3af622322a7e0386293d6d58427ba5ad7438af218f43fba948e981aeda736f72

Observation 63524a39-12b4-465e-b17e-d1473aab0199 · outbound

This paper cites Lacker and T.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Lacker and T

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.569224Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:dfc3780fcff2e35f56a9a2251a05e1698e2164d284dc4eb247a1a3f4a8a7a2a7

Observation f742a5d8-82e6-4acf-9401-eea3433589c3 · outbound

This paper cites Lasry and P.L.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Lasry and P.L

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.551141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:2049e01c92baaedab7b3cd61c1d1ebf2f730f532b34979aceed05ceef7781371

Observation 1de2893c-0493-4b99-ad28-d8c28a704ee9 · outbound

This paper cites Learning in Mean Field Games: A Survey.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Learning in Mean Field Games: A Survey

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.947857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:4b4149ace34ad09d36887de5702c5dc247cbee615640e64ab404817ceed72abc

Observation b8747aaa-1dfa-41e6-b75f-7b9a0210cfb2 · outbound

This paper cites Liang, Z.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Liang, Z

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.534454Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:de40f0d85259ec17ed35616d2e050046787db854ae984c1d647ec90d7f33eb27

Observation b0c7817d-5ec0-4d90-98cd-27e43621de98 · outbound

This paper cites Lions (2006): Cours au coll\` e ge de france: Th\' e orie des jeux \` a champ moyens.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Lions (2006): Cours au coll\` e ge de france: Th\' e orie des jeux \` a champ moyens

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.537835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:271dd9097eb6f98520047b41ca6e0359c9a7d0ec9071d859f85185b9107a583b

Observation ee388102-408a-49b7-99e2-cdc3ede0ef73 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.528870Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:0619fd59a2a3b32c0455624692a4e427775be1dc4beaa4b1ecca2da90718ff3f

Observation fe040c8f-c729-4115-ac52-abbd5f3b2eab · outbound

This paper cites Motte and H.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Motte and H

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.525766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:472a0e81978b10e1b3aae14b93c80e804c6c8f8fc94ca083cde7187fea609f45

Observation dcdc3bd5-a2ab-4ee9-8614-d590f42e5d2c · outbound

This paper cites Mondal, M.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Mondal, M

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.531774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:e0ab6fcb48b8748c712cc1ae6daf1d1b3ded5538f5a7f439582dafaadfdbbc22

Observation 52b2e4b8-a24a-4de8-9698-051b9470f4c1 · outbound

This paper cites Mean-Field Control based Approximation of Multi-Agent Reinforcement Learning in Presence of a Non-decomposable Shared Global State.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Mean-Field Control based Approximation of Multi-Agent Reinforcement Learning in Presence of a Non-decomposable Shared Global State

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.925565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:c4779416945f1f3df212e525b0aefde5e00bb18446ce820d1aec4e79689a1e02

Observation fcf30e3c-e6aa-498d-940a-d66fe3a450cc · outbound

This paper cites Efficient Model-Based Multi-Agent Mean-Field Reinforcement Learning.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Efficient Model-Based Multi-Agent Mean-Field Reinforcement Learning

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-12T09:56:27.921031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:280f87d46c592053dc30c16792623495bc9bccf85d85d3ffd6629793c6fd38cf

Observation e2ef8d98-fb88-444e-8949-6f3a50f014c5 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 53

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.541495Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:b525aa8ee236c0699cdeddec74172aa24569a4838524f3771b3775477edcf1b3

Observation 4f941bbb-5569-4e42-9690-2c756fdad3b9 · outbound

This paper cites Pham and X.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Pham and X

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.547329Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:b20d9afb8a545baea98f2b7fecd3887ba22e3f2f3c27a863154099a277de8262

Observation d81143eb-ae52-44c3-a9fa-9bba3b5bd041 · outbound

This paper cites Pham and X.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Pham and X

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.520617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:a477845be506c3bfda117fb4187ec4ab80f225070feb4e9889e7a52627be8b10

Observation 4f3673fc-f1c1-466b-898b-17d7eb7df07a · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 56

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.523168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:1ab7873ed929dffe6f134ef41f4f8e63fb2eb18b5162240a6cbf3063b0e9e53a

Observation 7717e50b-1166-402b-b264-8fd26f8d30f1 · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.517830Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:69cfb45eb3c0e45a86b4965b0b3b2c7c1d0a301f1e22946eadca991bfd7053ce

Observation fa99ffa9-d6ca-4966-b63b-f1c4b6a6e085 · outbound

This paper cites Gao and L.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Gao and L

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.554381Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:d478b09f4331462d6d315aab6be652c6a668f221daa794678302f4d5b25c96e8

Observation ffed49b6-2f42-4657-b533-c360afc5ab1b · outbound

This paper cites an unresolved cited work.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unresolved cited work

Reference 59

Resolution
unresolved
raw_fallback, observed 2026-05-27T08:49:00.544492Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:b6338680f42a4c13d26956de1da4fb2ef99390c32f7fe5608ef8a3115e5bc4f9

Observation 1918b7d7-f3ac-4919-ada7-7c71efa2cef7 · outbound

This paper cites Watkins and P.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Watkins and P

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.512590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:8cb2e42bf6e392c0bdfd525bf4d8f61e16c1f669a284f2dd547a93792c709c39

Observation a692dbe6-2b29-415c-8da0-3c67b7c6a987 · outbound

This paper cites Wei and X.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Wei and X

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-05-27T08:49:00.509825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:278c3e2841209a1d40703ef0a44811e8bbc8c0bb599ffdef26d90d393a3def95

Observation af00bf24-b6b7-4ed5-8720-a9996e0f244a · outbound

This paper cites Unified continuous-time q-learning for mean-field game and mean-field control problems.

Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms Unified continuous-time q-learning for mean-field game and mean-field control problems

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-08-03T01:10:22.407760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=arxiv_source observed=2026-05-07T08:41:31.247249Z digest=sha256:55be187b8da136ea93051eba2bb9a323b658379059556cc3f8b97c116215e385

Pith citing papers

Observation 6f28834f-d3fc-4abe-b1c3-6319079a036a · inbound

Robust $Q$-learning for mean-field control under Wasserstein uncertainty in common noise cites this paper.

Robust $Q$-learning for mean-field control under Wasserstein uncertainty in common noise Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms

Reference 73

Resolution
verified exact
local_arxiv, observed 2026-07-04T05:29:35.407820Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-04T06:34:03.388597+00:00.

source=pdf_text observed=2026-06-26T15:59:25.061726Z digest=sha256:e46bab7cdabbba9c86dff74dbbd71c95991b5b02699e15c2f0363bcf8342bc33

Observation 3a4d0c93-e8b0-4a88-91fa-cce5ffa2ac22 · inbound

Actor-Critic Learning for Extended Mean Field Control with Deterministic Policies cites this paper.

Actor-Critic Learning for Extended Mean Field Control with Deterministic Policies Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-14T07:42:38.840826Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T07:42:38.840826Z digest=sha256:f0be6d70b9a4ae3b6d58a76f9e5e1d39f1be505a45cc723042d0370f99659ff4