English

Actor critic learning algorithms for mean-field control with moment neural networks

Machine Learning 2023-09-11 v1 Machine Learning Optimization and Control

Abstract

We develop a new policy gradient and actor-critic algorithm for solving mean-field control problems within a continuous time reinforcement learning setting. Our approach leverages a gradient-based representation of the value function, employing parametrized randomized policies. The learning for both the actor (policy) and critic (value function) is facilitated by a class of moment neural network functions on the Wasserstein space of probability measures, and the key feature is to sample directly trajectories of distributions. A central challenge addressed in this study pertains to the computational treatment of an operator specific to the mean-field framework. To illustrate the effectiveness of our methods, we provide a comprehensive set of numerical results. These encompass diverse examples, including multi-dimensional settings and nonlinear quadratic mean-field control problems with controlled volatility.

Keywords

Cite

@article{arxiv.2309.04317,
  title  = {Actor critic learning algorithms for mean-field control with moment neural networks},
  author = {Huyên Pham and Xavier Warin},
  journal= {arXiv preprint arXiv:2309.04317},
  year   = {2023}
}

Comments

16 pages, 11 figures

R2 v1 2026-06-28T12:16:14.460Z