Pith. sign in

REVIEW

Actor critic learning algorithms for mean-field control with moment neural networks

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2309.04317 v1 pith:N3CZULAX submitted 2023-09-08 stat.ML cs.LGmath.OC

classification stat.MLcs.LGmath.OC
keywords mean-fieldcontrollearningactorcriticfunctionmomentneural
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
read the original abstract

We develop a new policy gradient and actor-critic algorithm for solving mean-field control problems within a continuous time reinforcement learning setting. Our approach leverages a gradient-based representation of the value function, employing parametrized randomized policies. The learning for both the actor (policy) and critic (value function) is facilitated by a class of moment neural network functions on the Wasserstein space of probability measures, and the key feature is to sample directly trajectories of distributions. A central challenge addressed in this study pertains to the computational treatment of an operator specific to the mean-field framework. To illustrate the effectiveness of our methods, we provide a comprehensive set of numerical results. These encompass diverse examples, including multi-dimensional settings and nonlinear quadratic mean-field control problems with controlled volatility.

Discussion (0). Continue with ORCID to comment.

Pith tools