Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:10:27.705945Z
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 99 of 99 outbound references and 1 inbound Pith citation observation for arXiv:2505.22781.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:10:27.705945Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-05-13T01:17:41.135172Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-13T01:27:02.760603Z
99 of 99 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation bf637fce-5932-47bb-a84d-fccde22410ae · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Capuzzo-Dolcetta, I
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4a6670e5-2aa5-40cc-b372-5d783f349faa · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Porretta, A
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 453d78a0-a8be-4b55-9a06-4050d75d6c03 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean field games: numerical methods for the planning problem
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ccbebb40-1486-421b-adef-e6689a6eff64 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean field games and applications: Numerical aspects
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 50fb8c21-04b9-48c5-9ee9-c258f2db1281 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99b8884a-b888-4366-8840-6503594e9d34 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games An extended mean field game for storage in smart grids
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c2ea739-69c9-4425-b6bb-f1b752ee86ad · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Regularization of the policy updates for stabilizing mean field games
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68a8815d-4929-4184-b623-1ef5e7930c55 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games D., and Saldi, N
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5f18d09d-74d5-4ad8-aa92-cd8971e47d3c · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games D., and Saldi, N
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29ad1f48-6d67-45d6-8117-15ccfb0cfe63 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean-field sampling for cooperative multi-agent reinforcement learning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 570de5b3-c6e6-4ed4-84d7-b2cd89a7bb11 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unified Reinforcement Q-Learning for Mean Field Game and Control Problems
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a115e582-f4c1-4f97-874b-d709fb4e066e · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Convergence of Multi-Scale Reinforcement Q-Learning Algorithms for Mean Field Game and Control Problems
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d9faeb28-e1ab-46f7-9b5b-4d91e7518fbd · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games On solutions of mean field games with ergodic cost
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c38864c-386d-4df4-b653-44aafd25fbe2 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Lipschitz continuity in model-based reinforcement learning
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d1c12000-e29d-4a60-afc9-9eed044892e0 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Inapproximability of np-complete variants of nash equilibrium
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb65bbb7-a39b-440f-8e8e-874cfdae2ee1 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games M., Munos, R., and Kappen, H
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d527bfc3-31dc-47f3-9bcb-745d3b78ebf6 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Priuli, F
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0bb83870-e706-48a0-8f24-a83d215f526a · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A mean-field game model of electricity market dynamics
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a13ea33-c593-4ce7-b35b-3ddd47b16e40 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Hesse, S
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7d1e2b1d-41c2-481a-9346-d99ba8da6420 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games First-order methods in optimization
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be756859-2ab8-4cf2-a5ca-494054b2bbe8 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Russo, D
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48e4c65e-cc51-4081-b65b-5e4d78f9ff8f · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A., Ortega, P
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6bd91a86-fb64-40ac-8366-1dfc1730f3f9 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games The master equation and the convergence problem in mean field games:(ams-201)
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7dce5b72-226e-41ef-bbc0-72f17a6825ad · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Lauri \`e re, M
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9096057d-4782-48ed-a291-3e6a5577120b · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean Field Games and Systemic Risk
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation db6ebada-1185-4011-b4e7-f9bebad60ca4 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Probabilistic theory of mean field games with applications I-II
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 25424109-0fc2-4849-905e-af23cd3e2c59 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Numerical method for fbsdes of mckean--vlasov type
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6b9d0135-894c-417e-83a5-c4d104926724 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e93f1e8-f4a9-473f-a040-9bb9bc1dec24 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Koeppl, H
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 27bbb1b0-3aba-4b92-9e01-73ad00a0db6d · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Koeppl, H
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dec370c1-d457-45b3-bb97-7d9744f31cb0 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A mean field game analysis of sir dynamics with vaccination
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dd77e7cb-60ec-45b4-9aab-1a478328e33f · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Touzi, N
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation f47c2e38-69af-4291-97e3-dfd0e973e66b · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Silva, F
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 97469166-6922-49d8-8dd6-e7e430524278 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games N-player games and mean field games of moderate interactions
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 023d22f7-2603-4386-8fbe-f1dde2ae9152 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Counterfactual multi-agent policy gradients
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 79bbc963-3836-4e96-9a35-0f8520672c3c · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Convergence of adaptive and interacting markov chain monte carlo algorithms
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 03d71b27-7be9-4f46-a629-b5e3f2106beb · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Taming the noise in reinforcement learning via soft updates
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b1b27f0a-1f7f-4f2d-9c0f-95401d9ff0e1 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A theory of regularized markov decision processes
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78fb16c5-db85-4513-951f-93cf8e0bbccb · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Concave Utility Reinforcement Learning: the Mean-Field Game Viewpoint
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d69e59ea-a95b-4796-8bf2-d7fbc8088ae4 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Numerical resolution of mckean-vlasov fbsdes using neural networks
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 3a1a6c36-559e-48d7-94c7-a7e450b1a6bf · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Diepold, K
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5bfd2c8a-fd34-4a50-ae8f-e06c4b1236c4 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Learning mean-field games
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation ccd5ec80-907a-4e95-8d0e-98e07be6c675 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A general framework for learning mean-field games
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e65f3309-8d85-4952-bc78-5a6fae295914 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Reinforcement learning with deep energy-based policies
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 5b074396-72b5-4276-ba77-c170a92c8aa2 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games J., Liaw, C., Plan, Y., and Randhawa, S
Reference 45
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 94f56b99-100a-4721-8c09-56180e6a5ef6 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8a7b82f-be21-40f2-bce0-cb5781e8892a · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0a887376-c623-4d30-a983-83a98eab0cb3 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games E., and Malham \'e , R
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 18f08f07-8014-474d-ad13-e41ec9bc9bf9 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games P., and Caines, P
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 39a6dc44-f5cd-4cd2-a412-8b19c3e48907 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games P., and Caines, P
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation db69e7c9-770d-46b8-9d34-f9b69d01db74 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games P., and Caines, P
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 0868478e-fd85-46bb-9d00-5a1b1ab40a13 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Sha, F
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9f637947-4b9e-4903-a1cf-004bc8ed7e5d · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Langford, J
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d2cf9ab8-375c-4972-aa89-067c533187b7 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation cbab3140-898f-4cc1-af1f-5adf410ed2c3 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Singh, S
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a733ded7-9ecf-4b7d-a69f-896fc7515207 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A., and Peters, J
Reference 56
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b216d78f-32db-4a42-9f27-166e20c2548e · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Zariphopoulou, T
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6137cd43-9af3-45e5-b098-e70a40ca2750 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Lions, P.-L
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 00e46286-6eca-46ee-98e3-ab074020cf14 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Lions, P.-L
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation dc7ff6ff-efdf-4703-ab72-e57f9deb2496 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Lions, P.-L
Reference 60
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation df7cf107-2d2d-46dd-ad91-acc63b4c49c2 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Learning in Mean Field Games: A Survey
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caa7464b-ca39-4d2a-b1a3-52c81f97e0b2 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Tankov, P
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4ef1a75-91d8-4d15-95cf-65c2afa50d05 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games G., and Castro, P
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 061924ee-e8d9-42d4-a570-edba6260118a · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A mean-field game approach to cloud resource management with function approximation
Reference 64
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e4d9343c-f81b-4127-ac4e-b3612d920a9b · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games I., Fern \'a ndez-Gaucherand, E., Hern \'a ndez-Hernandez, D., Coraluppi, S., and Fard, P
Reference 65
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 62b0348e-6afa-4d28-a1a0-b1e1d296efc5 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games J., and Le Fort-Piat, N
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c2192078-c026-4f09-9341-aeb0ddfb1cd0 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games On the global convergence rates of softmax policy gradient methods
Reference 67
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 179377ee-779a-40c0-adfe-684e4f89a49c · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Asynchronous methods for deep reinforcement learning
Reference 68
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation deee8e41-6131-4519-83d0-0d81307edf98 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Equilibrium points in n-person games
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9459171a-5cd1-4b5f-a64c-ace63906c4e2 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Non-cooperative games
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8797b717-c186-4857-a0cb-d49e42ae35a9 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games A unified view of entropy-regularized Markov decision processes
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d45e711e-b08b-4bf2-b8aa-78fad369ea65 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Combining policy gradient and q-learning
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6a6656f1-46f9-4cf2-a0f6-3db7528d17c0 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Scaling mean field games by online mirror descent
Reference 73
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d341e81a-6d43-4e45-9261-cf9318c72e8f · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Fictitious play for mean field games: C ontinuous time analysis and applications
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 9804e0d7-ba49-4ec8-b59b-1efd1be482c1 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Mean Field Games Flock! The Reinforcement Learning Way
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed45f950-7f0f-4a09-9c65-7cf997065bfc · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Generalization in mean field games by learning master policies
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 6f072f44-c0b5-4a69-907c-fc5f15848068 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Relative entropy policy search
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a6d8c63e-2826-4cb1-8ef4-c27bdf46da08 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation b27f4cb1-4c6d-4247-9f29-14222d27e36e · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Risk-averse dynamic programming for markov decision processes
Reference 79
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 78eb7089-8b69-4d91-8dea-1905199a3f2a · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Discrete-time average-cost mean-field games on polish spaces
Reference 80
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a06817dc-94c0-40e9-b3b1-aced4a8a7c3e · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games The StarCraft Multi-Agent Challenge
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27eef3f0-dde0-4a7e-9fbb-adfc6193a832 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Geist, M
Reference 82
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 17cde2f0-fc39-4990-83b9-63c2af12ed7f · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Trust region policy optimization
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e163f9de-fd01-4a78-be95-55c051baff8d · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Safe, Multi-Agent, Reinforcement Learning for Autonomous Driving
Reference 84
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62a70fbc-ee7d-4dfb-82e9-fedcc03f6342 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Adaptive trust region policy optimization: Global convergence and faster rates for regularized mdps
Reference 85
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c61aa626-fc8a-479d-9489-b171f6d3fac2 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Near-optimal time and sample complexities for solving markov decision processes with a generative model
Reference 86
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c285cf1f-31d3-422b-9690-aae4aa618c9a · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Barto, A
Reference 87
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation bd5dc256-aa0c-4404-8f4f-3f0e0ae72acd · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 88
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 8f602fc9-608c-4b01-8d32-15962ae3e3ba · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Algorithms for reinforcement learning
Reference 89
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d97a3575-cbc2-4494-a256-039da6b3be71 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games and Zhou, X
Reference 90
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation d23c640d-f2cb-4332-a07e-46bc412993b2 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Deep learning-based numerical methods for high-dimensional parabolic partial differential equations and backward stochastic differential equations
Reference 91
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 14ce55e6-3de5-4025-be1d-644c4611101b · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 92
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation c4948045-4cdc-45ec-9ad2-856d0e229c0e · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 93
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 97283845-d13d-4399-bf3c-36ce0d4ffa19 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Policy mirror ascent for efficient and independent learning in mean field games
Reference 94
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation e4890fa3-9eba-4808-83e8-edca356f2b54 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 95
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 83382493-16e4-41c4-bcf8-b5ca657e6f61 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Multi-agent reinforcement learning: A selective overview of theories and algorithms
Reference 96
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation a2658e27-0b40-4d5f-b58c-a51f579f4b39 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games Unresolved cited work
Reference 97
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 010bacfd-047a-4a62-9ea7-646781f58239 · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games D., Bagnell, J
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.
Observation 26a31921-354f-45b5-b123-62d2532794ca · outbound
Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games write newline
Reference 99
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d4fce0ae-be86-4a46-ab5c-2eb9ce8dd9b7 · inbound
Towards Model-Free Learning in Dynamic Population Games: An Application to Karma Economies Finite-Sample Convergence Bounds for Trust Region Policy Optimization in Mean-Field Games
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.