Pith. sign in

arXiv preprint arXiv:2312.06659 (2023)

2 Pith papers cite this work. Polarity classification is still indexing.

2 Pith papers citing it

fields

math.OC 2

years

2026 2

verdicts

UNVERDICTED 2

representative citing papers

Mean Field Reinforcement Learning

math.OC · 2026-07-01 · unverdicted · novelty 2.0

A monograph develops the probabilistic and control-theoretic framework connecting multi-agent reinforcement learning to mean field control, including analyses of Q-learning, policy gradients, and numerical methods for linear-quadratic and general models.

citing papers explorer

Showing 2 of 2 citing papers.

  • Continuous-time q-learning for mean-field control with common noise, part-II: q-learning algorithms math.OC · 2026-04-30 · unverdicted · none · ref 5

    The authors propose actor-critic q-learning algorithms for mean-field control with common noise based on martingale orthogonality conditions and relaxed controls, establish convergence of inner iterations in the linear-quadratic case, and demonstrate performance on examples.

  • Mean Field Reinforcement Learning math.OC · 2026-07-01 · unverdicted · none · ref 9

    A monograph develops the probabilistic and control-theoretic framework connecting multi-agent reinforcement learning to mean field control, including analyses of Q-learning, policy gradients, and numerical methods for linear-quadratic and general models.