Pith. sign in

Paper Citation Record · LEDGER

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games

As of 9 August 2026, this Paper Citation Record lists 99 of 99 outbound references and 1 inbound Pith citation observation for arXiv:2505.22781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.22781 v1

Coverage vector

measured 99 of 99 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:10:27.705945Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-05-13T01:17:41.135172Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-13T01:27:02.760603Z

Reference resolution

99 of 99 outbound references displayed

  • verified exact2
  • verified fuzzy54
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation bf637fce-5932-47bb-a84d-fccde22410ae · outbound

This paper cites and Capuzzo-Dolcetta, I.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Capuzzo-Dolcetta, I

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:12.196564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:12.196564Z digest=sha256:e59781bcc82ad2fe985fdf3e139c30b0a0eae95e06ad28c1d2361907364a0cf4

Observation 4a6670e5-2aa5-40cc-b372-5d783f349faa · outbound

This paper cites and Porretta, A.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Porretta, A

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:12.329482Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:12.329482Z digest=sha256:083913b70451116fa183c8b7504ae22c530c0a366a988218b2d4a982dec51a51

Observation 453d78a0-a8be-4b55-9a06-4050d75d6c03 · outbound

This paper cites Mean field games: numerical methods for the planning problem.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean field games: numerical methods for the planning problem

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:12.491661Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:12.491661Z digest=sha256:3e093bb7e1f89ddbacba40a9d0ed20d443756d03f1003a4bc060db7843e4d18f

Observation ccbebb40-1486-421b-adef-e6689a6eff64 · outbound

This paper cites Mean field games and applications: Numerical aspects.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean field games and applications: Numerical aspects

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:12.619752Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:12.619752Z digest=sha256:51987169be3809e00bd4aaedfd4799b4a108b9be7cb1947eeea6ac9a145fc1a4

Observation 50fb8c21-04b9-48c5-9ee9-c258f2db1281 · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:12.806388Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:12.806388Z digest=sha256:fcc3aaec843741b5046fc58e8cb3309fbb52c2734227119c585eb9ac8a57a923

Observation 99b8884a-b888-4366-8840-6503594e9d34 · outbound

This paper cites An extended mean field game for storage in smart grids.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games An extended mean field game for storage in smart grids

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:13.017801Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:13.017801Z digest=sha256:44a0899d4700c2931c01db6453d6859970bf425708d1956a335e92be46ca742d

Observation 5c2ea739-69c9-4425-b6bb-f1b752ee86ad · outbound

This paper cites Regularization of the policy updates for stabilizing mean field games.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Regularization of the policy updates for stabilizing mean field games

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:13.221565Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:13.221565Z digest=sha256:752891d3aed2e33107a96279d9d1355b7197003f3e7da1a6f570edec7e55b2dc

Observation 68a8815d-4929-4184-b623-1ef5e7930c55 · outbound

This paper cites D., and Saldi, N.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games D., and Saldi, N

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:13.373064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:13.373064Z digest=sha256:4c34b60368c4cbf553c9969d77900a47aa84236ecf2e6117ccdd8f0fa0b906d3

Observation 5f18d09d-74d5-4ad8-aa92-cd8971e47d3c · outbound

This paper cites D., and Saldi, N.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games D., and Saldi, N

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:13.491777Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:13.491777Z digest=sha256:957b7683ab457c6e6d607a2510a32775a07e4da64e835123cff28dc6273504d6

Observation 29ad1f48-6d67-45d6-8117-15ccfb0cfe63 · outbound

This paper cites Mean-field sampling for cooperative multi-agent reinforcement learning.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean-field sampling for cooperative multi-agent reinforcement learning

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:13.621544Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:13.621544Z digest=sha256:7c37bd849556d7b141dc511c1522b6e35f0e442bb2ab3444504d8c1d216108b1

Observation 570de5b3-c6e6-4ed4-84d7-b2cd89a7bb11 · outbound

This paper cites Unified Reinforcement Q-Learning for Mean Field Game and Control Problems.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unified Reinforcement Q-Learning for Mean Field Game and Control Problems

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:10:28.517786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:13.774546Z digest=sha256:edbcfe0f858a4783fabfcde07b4fd76cbf0763f1e54b424fca15142be18c6b2a

Observation a115e582-f4c1-4f97-874b-d709fb4e066e · outbound

This paper cites Convergence of Multi-Scale Reinforcement Q-Learning Algorithms for Mean Field Game and Control Problems.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Convergence of Multi-Scale Reinforcement Q-Learning Algorithms for Mean Field Game and Control Problems

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:13.932048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:13.932048Z digest=sha256:3dd49144f8e69989d21f90e3fc307223b0c7c9f76fb20e19af38eefc1901e14e

Observation d9faeb28-e1ab-46f7-9b5b-4d91e7518fbd · outbound

This paper cites On solutions of mean field games with ergodic cost.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games On solutions of mean field games with ergodic cost

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:14.074540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:14.074540Z digest=sha256:7b9330529f3e9c0948f7acae60395b3f5f1409c42411ab97ab3ca898e720bdcc

Observation 7c38864c-386d-4df4-b653-44aafd25fbe2 · outbound

This paper cites Lipschitz continuity in model-based reinforcement learning.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Lipschitz continuity in model-based reinforcement learning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:14.243790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:14.243790Z digest=sha256:8d91251da7387d26dacb78679062daeba0b9bcaa4d03340e5f85ae3a747608cf

Observation d1c12000-e29d-4a60-afc9-9eed044892e0 · outbound

This paper cites Inapproximability of np-complete variants of nash equilibrium.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Inapproximability of np-complete variants of nash equilibrium

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:14.377895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:14.377895Z digest=sha256:07477e1b22e15b121b1125856b936d95a49f4a53af8aa740693025868bfc0d71

Observation cb65bbb7-a39b-440f-8e8e-874cfdae2ee1 · outbound

This paper cites M., Munos, R., and Kappen, H.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games M., Munos, R., and Kappen, H

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:14.540084Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:14.540084Z digest=sha256:0d61d9774996956665ebc365558ee5ab2e54ade06327149cd3c18cefda5eb43c

Observation d527bfc3-31dc-47f3-9bcb-745d3b78ebf6 · outbound

This paper cites and Priuli, F.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Priuli, F

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:14.688053Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:14.688053Z digest=sha256:582b0b971178db658e38ae0cf26e382af23a89d816f555d39412b587c7858cec

Observation 0bb83870-e706-48a0-8f24-a83d215f526a · outbound

This paper cites A mean-field game model of electricity market dynamics.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A mean-field game model of electricity market dynamics

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:14.886881Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:14.886881Z digest=sha256:e68459a62d0817a61197d7f47e731a96c83e3e02f43a54795a2cb40262fca1de

Observation 7a13ea33-c593-4ce7-b35b-3ddd47b16e40 · outbound

This paper cites and Hesse, S.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Hesse, S

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:15.037956Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:15.037956Z digest=sha256:76e3d1fa4cfd0349c3eb583676f4f797afd95235cec4ad83c2f91ca7b2bc368e

Observation 7d1e2b1d-41c2-481a-9346-d99ba8da6420 · outbound

This paper cites First-order methods in optimization.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games First-order methods in optimization

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:15.203972Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:15.203972Z digest=sha256:30542a93da86bc803a0b07e6e75a91ddd1c6287415bfebde477b9c54c042d34b

Observation be756859-2ab8-4cf2-a5ca-494054b2bbe8 · outbound

This paper cites and Russo, D.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Russo, D

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:15.404955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:15.404955Z digest=sha256:c5267005a1dca5209e46d223ec89f2a09f8dd8ae15ac05965d31bd6812820e1a

Observation 48e4c65e-cc51-4081-b65b-5e4d78f9ff8f · outbound

This paper cites A., Ortega, P.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A., Ortega, P

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:15.613780Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:15.613780Z digest=sha256:95f28d5b2434a1d0bc3046f85f328340a4e7de6a6b45e4f56f79350400762ede

Observation 6bd91a86-fb64-40ac-8366-1dfc1730f3f9 · outbound

This paper cites The master equation and the convergence problem in mean field games:(ams-201).

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games The master equation and the convergence problem in mean field games:(ams-201)

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:15.729213Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:15.729213Z digest=sha256:3cfe437dc494c52110d72d7ea30def4dd06e67f42641908f9a200a335e0798ce

Observation 7dce5b72-226e-41ef-bbc0-72f17a6825ad · outbound

This paper cites and Lauri \`e re, M.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Lauri \`e re, M

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:45.706097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:15.885470Z digest=sha256:5de2ac789773ecfc4e6ff0bd70ce2802287c452cffbb2359aa5a6d7e6aaf1fbe

Observation 9096057d-4782-48ed-a291-3e6a5577120b · outbound

This paper cites Mean Field Games and Systemic Risk.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean Field Games and Systemic Risk

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:10:28.223000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:16.063803Z digest=sha256:3aa72f740c852eeeaa4de1d7d929a09ba13c583fd0b1a8589fa10d6a3762ed7e

Observation db6ebada-1185-4011-b4e7-f9bebad60ca4 · outbound

This paper cites Probabilistic theory of mean field games with applications I-II.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Probabilistic theory of mean field games with applications I-II

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:45.413359Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:16.251055Z digest=sha256:c31cc839f20a35c8aa28975a47c734e45aec54bc4f2acb1a24d7178147d41a1f

Observation 25424109-0fc2-4849-905e-af23cd3e2c59 · outbound

This paper cites Numerical method for fbsdes of mckean--vlasov type.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Numerical method for fbsdes of mckean--vlasov type

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:45.138057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:16.373548Z digest=sha256:a5386d6422859317dbf04f2d9020d7918e88e2726ca2a6d52cc7f8c2309db224

Observation 6b9d0135-894c-417e-83a5-c4d104926724 · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:16.529488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:16.529488Z digest=sha256:f2b0792fd0c53504efadf6c5c359bf0920ab818ee07d5be114752191ec1b2ed1

Observation 2e93f1e8-f4a9-473f-a040-9bb9bc1dec24 · outbound

This paper cites and Koeppl, H.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Koeppl, H

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:44.865814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:16.744269Z digest=sha256:dd848efa484add40e120575769493cd2c9d311a8baf87cd46b29b540fd53398c

Observation 27bbb1b0-3aba-4b92-9e01-73ad00a0db6d · outbound

This paper cites and Koeppl, H.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Koeppl, H

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:44.565375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:16.855232Z digest=sha256:624401ecaa8680a607fb77a4384a6334e126db0857790604a53f35aeb4fe6902

Observation dec370c1-d457-45b3-bb97-7d9744f31cb0 · outbound

This paper cites A mean field game analysis of sir dynamics with vaccination.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A mean field game analysis of sir dynamics with vaccination

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:44.298047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:17.038129Z digest=sha256:08dc786dc3cccc24d7cd8c1be6b9d844b48ae0107c7e244916165478512a8afd

Observation dd77e7cb-60ec-45b4-9aab-1a478328e33f · outbound

This paper cites and Touzi, N.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Touzi, N

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:43.918293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:17.232176Z digest=sha256:921e0c163d0611f8d29fc6dc2b3576c7246ac326c2bb9d4684ad3bf20fd1935d

Observation f47c2e38-69af-4291-97e3-dfd0e973e66b · outbound

This paper cites and Silva, F.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Silva, F

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:43.646217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:17.431580Z digest=sha256:175b09206342ae6d1ac18bb79e2cb61f2df9a8df68819babfeddf211b478a03b

Observation 97469166-6922-49d8-8dd6-e7e430524278 · outbound

This paper cites N-player games and mean field games of moderate interactions.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games N-player games and mean field games of moderate interactions

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:43.356158Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:17.587562Z digest=sha256:c5af6ced197aee716a154b90e7363d0f4dee786cb70d2d5c813b5295111b9edb

Observation 023d22f7-2603-4386-8fbe-f1dde2ae9152 · outbound

This paper cites Counterfactual multi-agent policy gradients.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Counterfactual multi-agent policy gradients

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:43.097433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:17.719823Z digest=sha256:2094cc2fd617c456a3bfd1db1cb520dcf87092c851103f3b21f8e63c9a4160eb

Observation 79bbc963-3836-4e96-9a35-0f8520672c3c · outbound

This paper cites Convergence of adaptive and interacting markov chain monte carlo algorithms.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Convergence of adaptive and interacting markov chain monte carlo algorithms

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:42.793131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:17.855496Z digest=sha256:3ed3db8edf704a4902cfe21d0d6395706c57a83f8af6d05f54ddb378170c1b25

Observation 03d71b27-7be9-4f46-a629-b5e3f2106beb · outbound

This paper cites Taming the noise in reinforcement learning via soft updates.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Taming the noise in reinforcement learning via soft updates

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:42.494859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:18.009623Z digest=sha256:23dc480e9b880227aa347e49d2729b837c43d2b54cf126e61f2057a41ee3a769

Observation b1b27f0a-1f7f-4f2d-9c0f-95401d9ff0e1 · outbound

This paper cites A theory of regularized markov decision processes.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A theory of regularized markov decision processes

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:18.193450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:18.193450Z digest=sha256:9f1e682f6ea819c6172d08fcbc7420de8d8bd09f168ed3773bf599184046d5b3

Observation 78fb16c5-db85-4513-951f-93cf8e0bbccb · outbound

This paper cites Concave Utility Reinforcement Learning: the Mean-Field Game Viewpoint.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Concave Utility Reinforcement Learning: the Mean-Field Game Viewpoint

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:18.346868Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:18.346868Z digest=sha256:d4c02233a169632bec02a0d0b7f15ed33584135418d645d552268e13cdbfff9d

Observation d69e59ea-a95b-4796-8bf2-d7fbc8088ae4 · outbound

This paper cites Numerical resolution of mckean-vlasov fbsdes using neural networks.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Numerical resolution of mckean-vlasov fbsdes using neural networks

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:42.256776Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:18.546112Z digest=sha256:350ea05d8d96f24727b37985a9302d40358c7fa9621fe51b86a2d353dd21161d

Observation 3a1a6c36-559e-48d7-94c7-a7e450b1a6bf · outbound

This paper cites and Diepold, K.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Diepold, K

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:18.689464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:18.689464Z digest=sha256:295de81571772d7d22f46fe2f5ddacfa89cea65558f2e326d18cb08fa810272a

Observation 5bfd2c8a-fd34-4a50-ae8f-e06c4b1236c4 · outbound

This paper cites Learning mean-field games.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Learning mean-field games

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:41.972833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:18.864106Z digest=sha256:917258df71a5ed8ee095a0315e61a4335c7c8ceac556a3d28358c74648bc7d00

Observation ccd5ec80-907a-4e95-8d0e-98e07be6c675 · outbound

This paper cites A general framework for learning mean-field games.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A general framework for learning mean-field games

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:41.639343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:19.028275Z digest=sha256:c50900420dd527b40a4fa36f104d67428725ab1db8122e3c4743e49b697676f9

Observation e65f3309-8d85-4952-bc78-5a6fae295914 · outbound

This paper cites Reinforcement learning with deep energy-based policies.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Reinforcement learning with deep energy-based policies

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:41.336299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:19.165381Z digest=sha256:08656126cafb9c76379ff9e5382158b7bdbaaa44b63c88b0c3a1360ed7481c66

Observation 5b074396-72b5-4276-ba77-c170a92c8aa2 · outbound

This paper cites J., Liaw, C., Plan, Y., and Randhawa, S.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games J., Liaw, C., Plan, Y., and Randhawa, S

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:40.989906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:19.339607Z digest=sha256:6fc2bd27af7dd13c98dde3ba70539814dfc1ee158c4e182c65d9644c276cab99

Observation 94f56b99-100a-4721-8c09-56180e6a5ef6 · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:19.489673Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:19.489673Z digest=sha256:277e5aebd1adddf198eabdebcb222afa19f6b1bcc2c85b49a9710f903f8b0280

Observation f8a7b82f-be21-40f2-bce0-cb5781e8892a · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:40.725380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:19.640667Z digest=sha256:2666a0b628a0513f4aa854b5c98cdb362100ae10d3c783d4b08fbaf943babfde

Observation 0a887376-c623-4d30-a983-83a98eab0cb3 · outbound

This paper cites E., and Malham \'e , R.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games E., and Malham \'e , R

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:40.511051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:19.736199Z digest=sha256:1d4994f75c3eaa5de2e1f5b4197842207f565b6cbf7cd4b132aa2bbb1cb690fc

Observation 18f08f07-8014-474d-ad13-e41ec9bc9bf9 · outbound

This paper cites P., and Caines, P.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games P., and Caines, P

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:40.194629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:19.880740Z digest=sha256:81b53101126c90dfe07b28291ea2e34f563308973eb7437886c06e76922cc0ea

Observation 39a6dc44-f5cd-4cd2-a412-8b19c3e48907 · outbound

This paper cites P., and Caines, P.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games P., and Caines, P

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:39.850916Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:20.085274Z digest=sha256:833416450b5b1344bde5dcdb49e4d91f57172c3021344c89288ddcb32c09a401

Observation db69e7c9-770d-46b8-9d34-f9b69d01db74 · outbound

This paper cites P., and Caines, P.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games P., and Caines, P

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:39.591521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:20.238426Z digest=sha256:f913653a3e6f02c98cf0e3d627cf463b6c7c8436914d434d64cb067ba1684223

Observation 0868478e-fd85-46bb-9d00-5a1b1ab40a13 · outbound

This paper cites and Sha, F.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Sha, F

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:39.225351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:20.405360Z digest=sha256:cf6d8e4c967c88f331b10e18066e8841e115f63df175dbf5a498d189ce4bcae9

Observation 9f637947-4b9e-4903-a1cf-004bc8ed7e5d · outbound

This paper cites and Langford, J.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Langford, J

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:38.952013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:20.536988Z digest=sha256:c904d31a1fd78b0d2977a5da37c11f7737a56eb76f2e0bc298dc1ad0e4884063

Observation d2cf9ab8-375c-4972-aa89-067c533187b7 · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:38.682202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:20.673691Z digest=sha256:ce5ea85177ac5d6a0b88683d8a48721c906983149a0a1e06b63c3ef07c06a11c

Observation cbab3140-898f-4cc1-af1f-5adf410ed2c3 · outbound

This paper cites and Singh, S.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Singh, S

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:38.365483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:20.845556Z digest=sha256:cc0d866fe38a413417f3d0bfd32f039c142b3e93722350f8ba380afbedad7616

Observation a733ded7-9ecf-4b7d-a69f-896fc7515207 · outbound

This paper cites A., and Peters, J.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A., and Peters, J

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:21.019331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:21.019331Z digest=sha256:7b10a9c22e3635eb518ea1954b158b6baa67a1bc8f2999a6958c143644089a40

Observation b216d78f-32db-4a42-9f27-166e20c2548e · outbound

This paper cites and Zariphopoulou, T.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Zariphopoulou, T

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:38.039128Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:21.140695Z digest=sha256:f29986b8f83418d862f0b60ab7aef3093e1a3c2ba9dea237747992fa00c0d467

Observation 6137cd43-9af3-45e5-b098-e70a40ca2750 · outbound

This paper cites and Lions, P.-L.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Lions, P.-L

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:37.682067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:21.319382Z digest=sha256:01b7d7864a2968eab36c1d4476d57d34ff1bcde4789cf75bf7fec1f1924eb7d4

Observation 00e46286-6eca-46ee-98e3-ab074020cf14 · outbound

This paper cites and Lions, P.-L.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Lions, P.-L

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:37.358005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:21.475442Z digest=sha256:9c035ad6566325f0c0d2f0789b768795d131855e11f70206c9d1004f97bf3516

Observation dc7ff6ff-efdf-4703-ab72-e57f9deb2496 · outbound

This paper cites and Lions, P.-L.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Lions, P.-L

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:37.110259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:21.678545Z digest=sha256:baae2bf130492618f12d39da825959df30500b2582d0a1f6287c295fb7a6d75b

Observation df7cf107-2d2d-46dd-ad91-acc63b4c49c2 · outbound

This paper cites Learning in Mean Field Games: A Survey.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Learning in Mean Field Games: A Survey

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:21.805809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:21.805809Z digest=sha256:274be21aa4717101ecc158ccf1d0957d5fc3202077f313960104be8e34b4bc2b

Observation caa7464b-ca39-4d2a-b1a3-52c81f97e0b2 · outbound

This paper cites and Tankov, P.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Tankov, P

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:21.939715Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:21.939715Z digest=sha256:428ba328b1dd0ad001351e52ec213db53338250fa999d16e5f03ff5512013ccf

Observation b4ef1a75-91d8-4d15-95cf-65c2afa50d05 · outbound

This paper cites G., and Castro, P.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games G., and Castro, P

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:36.778709Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:22.118395Z digest=sha256:4bff03790492a9c36779e2d3864d3b787e8b96078653b5e77426da0fd6d4b107

Observation 061924ee-e8d9-42d4-a570-edba6260118a · outbound

This paper cites A mean-field game approach to cloud resource management with function approximation.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A mean-field game approach to cloud resource management with function approximation

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:36.518742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:22.294882Z digest=sha256:aca0c45f16da5c79db81794a9abc54684800a3eea2a83d95d1ff0c716a891128

Observation e4d9343c-f81b-4127-ac4e-b3612d920a9b · outbound

This paper cites I., Fern \'a ndez-Gaucherand, E., Hern \'a ndez-Hernandez, D., Coraluppi, S., and Fard, P.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games I., Fern \'a ndez-Gaucherand, E., Hern \'a ndez-Hernandez, D., Coraluppi, S., and Fard, P

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:36.242074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:22.422162Z digest=sha256:5652f86ca8237e13d2c60775ae2804c422bb20afbc5b04edfb354c9bd7fb71e8

Observation 62b0348e-6afa-4d28-a1a0-b1e1d296efc5 · outbound

This paper cites J., and Le Fort-Piat, N.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games J., and Le Fort-Piat, N

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:35.938947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:22.542113Z digest=sha256:c27c1267b0e389b36823065be10a03a7289a9a9c7dfa1d30f832bc07933e840a

Observation c2192078-c026-4f09-9341-aeb0ddfb1cd0 · outbound

This paper cites On the global convergence rates of softmax policy gradient methods.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games On the global convergence rates of softmax policy gradient methods

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:35.674906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:22.629063Z digest=sha256:13420fcb2231f9c91ba25b0748a84d3f5d762b8e3bbb70bc8e5c23d2362c0fc0

Observation 179377ee-779a-40c0-adfe-684e4f89a49c · outbound

This paper cites Asynchronous methods for deep reinforcement learning.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Asynchronous methods for deep reinforcement learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:35.407930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:22.697117Z digest=sha256:34530003329358a1dfb3d8dbf5caa291011951332fec15c2dcaf21fa77e30d6e

Observation deee8e41-6131-4519-83d0-0d81307edf98 · outbound

This paper cites Equilibrium points in n-person games.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Equilibrium points in n-person games

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:35.017098Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:22.825724Z digest=sha256:f4e59246e6c235968824e866389932a63de9e8b8d438c5c6d979207288decca1

Observation 9459171a-5cd1-4b5f-a64c-ace63906c4e2 · outbound

This paper cites Non-cooperative games.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Non-cooperative games

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:34.652398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:22.945284Z digest=sha256:9a3f06d37de40bcd0cdd82e1c647d78c61cd63d4697c2b21e4035e3b2bd0ffc7

Observation 8797b717-c186-4857-a0cb-d49e42ae35a9 · outbound

This paper cites A unified view of entropy-regularized Markov decision processes.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A unified view of entropy-regularized Markov decision processes

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:23.087809Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:23.087809Z digest=sha256:acaa07234090bd3bb6438ef995418dd9bd6b64e7cfd827c73232046191c87db7

Observation d45e711e-b08b-4bf2-b8aa-78fad369ea65 · outbound

This paper cites Combining policy gradient and q-learning.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Combining policy gradient and q-learning

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:34.305505Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:23.258768Z digest=sha256:d4b2b5ec245657bdebc59d4381022648f9f5d279046ada483e31d23b8b39ae76

Observation 6a6656f1-46f9-4cf2-a0f6-3db7528d17c0 · outbound

This paper cites Scaling mean field games by online mirror descent.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Scaling mean field games by online mirror descent

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:33.964017Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:23.407117Z digest=sha256:ba2749e9981f0a370c92b62c12dda5edd645c53c8822b5014f70a8187eba91d5

Observation d341e81a-6d43-4e45-9261-cf9318c72e8f · outbound

This paper cites Fictitious play for mean field games: C ontinuous time analysis and applications.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Fictitious play for mean field games: C ontinuous time analysis and applications

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:33.667286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:23.669181Z digest=sha256:4abd8c7771aa935c8cecaaf884eb482aa2c174c47b394dbaeb4dc928b99f48ee

Observation 9804e0d7-ba49-4ec8-b59b-1efd1be482c1 · outbound

This paper cites Mean Field Games Flock! The Reinforcement Learning Way.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean Field Games Flock! The Reinforcement Learning Way

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:23.894935Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:23.894935Z digest=sha256:7b32d7f7c65406baa5032f7216eb46dc3170a3a2540610aa735965e2a443619e

Observation ed45f950-7f0f-4a09-9c65-7cf997065bfc · outbound

This paper cites Generalization in mean field games by learning master policies.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Generalization in mean field games by learning master policies

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:33.421759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:24.088295Z digest=sha256:4dbaa57109f33c6e2cee82343a62d25f58d8e8433476f6de5f140e13093e609a

Observation 6f072f44-c0b5-4a69-907c-fc5f15848068 · outbound

This paper cites Relative entropy policy search.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Relative entropy policy search

Reference 77

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:33.137548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:24.280777Z digest=sha256:a9e8f7a0e15d57958fa2d36acd06765f24e9a819561a5bb52ef0a84caaed8dcf

Observation a6d8c63e-2826-4cb1-8ef4-c27bdf46da08 · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 78

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:32.954863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:24.415989Z digest=sha256:88ad7eabf538a18304113a8f6ace47cb51710d3a1e091f7847ad0d3d51cbbd9c

Observation b27f4cb1-4c6d-4247-9f29-14222d27e36e · outbound

This paper cites Risk-averse dynamic programming for markov decision processes.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Risk-averse dynamic programming for markov decision processes

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:32.785228Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:24.501979Z digest=sha256:eb2acf95241cb1a1cdcbba9b99a521593376f9126a2377ee72d307f3f213a996

Observation 78eb7089-8b69-4d91-8dea-1905199a3f2a · outbound

This paper cites Discrete-time average-cost mean-field games on polish spaces.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Discrete-time average-cost mean-field games on polish spaces

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:32.563548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:24.662129Z digest=sha256:976220980f039ec4344c6cb8771da5da2ec1d2819a5e021709e42cd8f9379ffd

Observation a06817dc-94c0-40e9-b3b1-aced4a8a7c3e · outbound

This paper cites The StarCraft Multi-Agent Challenge.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games The StarCraft Multi-Agent Challenge

Reference 81

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:24.777756Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:24.777756Z digest=sha256:62eaf5811b7bf3cec76845e29132deb66d6683cd6a6e74bcbe96eef1e3bd1dbb

Observation 27eef3f0-dde0-4a7e-9fbb-adfc6193a832 · outbound

This paper cites and Geist, M.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Geist, M

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:32.297226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:24.904569Z digest=sha256:d872b63e64e8e965a7d04909085d6baf18116cb553b7eed5533b1881bec93ba3

Observation 17cde2f0-fc39-4990-83b9-63c2af12ed7f · outbound

This paper cites Trust region policy optimization.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Trust region policy optimization

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:32.042431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:25.085428Z digest=sha256:db612cf3b4e9ba46525ed9a48668462a0ba7fc548e117b38c8100114cd780fd0

Observation e163f9de-fd01-4a78-be95-55c051baff8d · outbound

This paper cites Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:25.197393Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:25.197393Z digest=sha256:2b41953b006c78a8c8eed53b923e8e205d86117925f77048cb4fd070fa8efb8f

Observation 62a70fbc-ee7d-4dfb-82e9-fedcc03f6342 · outbound

This paper cites Adaptive trust region policy optimization: Global convergence and faster rates for regularized mdps.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Adaptive trust region policy optimization: Global convergence and faster rates for regularized mdps

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:31.863235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:25.356684Z digest=sha256:01a0e60bc403efae3ba9f6726a2eb29dca18dbcc3c7b822a3c5aefce05d5f27f

Observation c61aa626-fc8a-479d-9489-b171f6d3fac2 · outbound

This paper cites Near-optimal time and sample complexities for solving markov decision processes with a generative model.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Near-optimal time and sample complexities for solving markov decision processes with a generative model

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:31.502983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:25.492393Z digest=sha256:009a254bb3f74c4be8bbebfd6297c18af616830630e20566e933e861fc3aaced

Observation c285cf1f-31d3-422b-9690-aae4aa618c9a · outbound

This paper cites and Barto, A.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Barto, A

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:31.332007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:25.663184Z digest=sha256:fd8e19b26fd13beda74d6e4045c30b6aa316afe0053ff9f1da8ba6f52798040f

Observation bd5dc256-aa0c-4404-8f4f-3f0e0ae72acd · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 88

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:31.146473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:25.789076Z digest=sha256:4cdec394a21ee2118adc6fae0bca583ceddb8a9c77321cc152b5c69decab6d3d

Observation 8f602fc9-608c-4b01-8d32-15962ae3e3ba · outbound

This paper cites Algorithms for reinforcement learning.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Algorithms for reinforcement learning

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:30.866770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:25.976783Z digest=sha256:7ac750ca41badebcb46a7a163f92eed968a432c1fdceb6f0d1985479091a1f1f

Observation d97a3575-cbc2-4494-a256-039da6b3be71 · outbound

This paper cites and Zhou, X.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Zhou, X

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:30.584759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:26.108732Z digest=sha256:f80b3d6ab0354952af8c2343a6a15333ff27637e4b003b0796b4bfff0b61cc8d

Observation d23c640d-f2cb-4332-a07e-46bc412993b2 · outbound

This paper cites Deep learning-based numerical methods for high-dimensional parabolic partial differential equations and backward stochastic differential equations.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Deep learning-based numerical methods for high-dimensional parabolic partial differential equations and backward stochastic differential equations

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:30.349052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:26.241320Z digest=sha256:16bc1fa776868aaa937fe297f3bfd0735b9d609c455a4e13c528fcb22b25bf47

Observation 14ce55e6-3de5-4025-be1d-644c4611101b · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 92

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:29.996593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:26.416453Z digest=sha256:ddbd34901d11a8efccf56cddf22b0d33bb772d79af2d20f7ac35774317050873

Observation c4948045-4cdc-45ec-9ad2-856d0e229c0e · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 93

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:29.720378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:26.554078Z digest=sha256:86300998b0b17de794564de5c82613672e85193f34596d84a8532ff270fe5a91

Observation 97283845-d13d-4399-bf3c-36ce0d4ffa19 · outbound

This paper cites Policy mirror ascent for efficient and independent learning in mean field games.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Policy mirror ascent for efficient and independent learning in mean field games

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:29.386478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:26.764903Z digest=sha256:f2091b45143a3e36c084ef53c78f454e6b82b300d8d7b178d4f30828950abf5c

Observation e4890fa3-9eba-4808-83e8-edca356f2b54 · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 95

Resolution
unresolved
raw_fallback, observed 2026-08-07T13:10:29.189716Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:27.017211Z digest=sha256:32debcb230084a7bb9573cd72dc2b252ae320ebaa92932f40e27a2bf697f1943

Observation 83382493-16e4-41c4-bcf8-b5ca657e6f61 · outbound

This paper cites Multi-agent reinforcement learning: A selective overview of theories and algorithms.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Multi-agent reinforcement learning: A selective overview of theories and algorithms

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:29.016638Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:27.275330Z digest=sha256:0191cf0088f5ba03a20ec0ad1a28810e7ab60e1f4c389382defc50a2da7008f1

Observation a2658e27-0b40-4d5f-b58c-a51f579f4b39 · outbound

This paper cites an unresolved cited work.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work

Reference 97

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:27.440566Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:27.440566Z digest=sha256:426433d89ff8e73f8b9be4bed71b8d7afd4435baa51a1f6a5ca9dcfc2875261a

Observation 010bacfd-047a-4a62-9ea7-646781f58239 · outbound

This paper cites D., Bagnell, J.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games D., Bagnell, J

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:10:28.811262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-07T13:10:27.525688Z digest=sha256:d28b81ba78395d56da038e31d69d3178c22095c47e01c461c869207002840d79

Observation 26a31921-354f-45b5-b123-62d2532794ca · outbound

This paper cites write newline.

Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games write newline

Reference 99

Resolution
unresolved
no resolver link, observed 2026-08-07T13:10:27.705945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T13:10:27.705945Z digest=sha256:d4ea16d413dfe56943482c46c46217191e0832cd842ba15ddfb89f402ec61bbb

Pith citing papers

Observation d4fce0ae-be86-4a46-ab5c-2eb9ce8dd9b7 · inbound

Towards Model-Free Learning in Dynamic Population Games: An Application to Karma Economies cites this paper.

Towards Model-Free Learning in Dynamic Population Games: An Application to Karma Economies Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-13T01:27:02.762569Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T01:17:41.135172Z digest=sha256:9697400e0a6ebbb4e2c5ecab785a8e0dab203a61eae619cc75effe86099d487c