English
Related papers

Related papers: Mean Field Markov Decision Processes

200 papers

In this paper, we consider the finite-state approximation of a discrete-time constrained Markov decision process (MDP) under the discounted and average cost criteria. Using the linear programming formulation of the constrained discounted…

Optimization and Control · Mathematics 2018-07-10 Naci Saldi

This paper studies the connections between mean-field games and the social welfare optimization problems. We consider a mean field game in functional spaces with a large population of agents, each of which seeks to minimize an individual…

Optimization and Control · Mathematics 2016-09-27 Sen Li , Wei Zhang , Lin Zhao

In this paper we model the role of a government of a large population as a mean field optimal control problem. Such control problems are constrainted by a PDE of continuity-type, governing the dynamics of the probability distribution of the…

Optimization and Control · Mathematics 2016-08-08 Giacomo Albi , Young-Pil Choi , Massimo Fornasier , Dante Kalise

In the Markov decision process model, policies are usually evaluated by expected cumulative rewards. As this decision criterion is not always suitable, we propose in this paper an algorithm for computing a policy optimal for the quantile…

Artificial Intelligence · Computer Science 2016-12-02 Hugo Gilbert , Paul Weng , Yan Xu

We investigate the problem of best policy identification in discounted linear Markov Decision Processes in the fixed confidence setting under a generative model. We first derive an instance-specific lower bound on the expected number of…

Machine Learning · Computer Science 2022-08-12 Jerome Taupin , Yassir Jedra , Alexandre Proutiere

We investigate mean-field games (MFG) in which agents can actively control their speed of access to information. Specifically, the agents can dynamically decide to obtain observations with reduced delay by accepting higher observation…

Optimization and Control · Mathematics 2025-06-03 Dirk Becherer , Christoph Reisinger , Jonathan Tam

We present a method for a certain class of Markov Decision Processes (MDPs) that can relate the optimal policy back to one or more reward sources in the environment. For a given initial state, without fully computing the value function,…

Machine Learning · Computer Science 2018-06-12 Josh Bertram , Peng Wei

In this paper, we study a class of discrete-time mean-field games under the infinite-horizon risk-sensitive discounted-cost optimality criterion. Risk-sensitivity is introduced for each agent (player) via an exponential utility function. In…

Optimization and Control · Mathematics 2018-10-08 Naci Saldi , Tamer Basar , Maxim Raginsky

In this article, we employ an input-output approach to expand the study of cooperative multi-agent control and optimization problems characterized by mean-field interactions that admit decentralized and selfish solutions. The setting…

Optimization and Control · Mathematics 2025-10-02 Vivek Khatana , Duo Wang , Petros Voulgaris , Nicola Elia , Naira Hovakimyan

We consider a discounted reward control problem in continuous time stochastic environment where the discount rate might be an unbounded function of the control process. We provide a set of general assumptions to ensure that there exists a…

Probability · Mathematics 2016-02-17 Dariusz Zawisza

This paper is devoted to studying constrained continuous-time Markov decision processes (MDPs) in the class of randomized policies depending on state histories. The transition rates may be unbounded, the reward and costs are admitted to be…

Probability · Mathematics 2012-01-04 Xianping Guo , Xinyuan Song

We study the optimal control of discrete time mean filed dynamical systems under partial observations. We express the global law of the filtered process as a controlled system with its own dynamics. Following a dynamic programming approach,…

Optimization and Control · Mathematics 2023-03-13 Jeremy Chichportich , Idris Kharroubi

We consider a control problem for a heterogeneous population composed of agents able to switch at any time between different options. The controller aims to maximize an average gain per time unit, supposing that the population is of…

Optimization and Control · Mathematics 2024-04-05 Quentin Jacquet , Wim van Ackooij , Clémence Alasseur , Stéphane Gaubert

We consider an optimal control problem where the average welfare of weakly interacting agents is of interest. We examine the mean-field control problem as the fluid approximation of the N-agent control problem with the setup of finite-state…

Optimization and Control · Mathematics 2024-02-13 Jingruo Sun

This paper introduces a new approach of treating platoon systems using mean-variance control formulation. The underlying system is a controlled switching diffusion in which the random switching process is a continuous-time Markov chain.…

Optimization and Control · Mathematics 2014-01-22 Zhixin Yang , G. Yin , Le Yi Wang , Hongwei Zhang

Markov decision processes (MDPs) are used to model a wide variety of applications ranging from game playing over robotics to finance. Their optimal policy typically maximizes the expected sum of rewards given at each step of the decision…

Machine Learning · Computer Science 2025-05-26 Maximilian Nägele , Jan Olle , Thomas Fösel , Remmy Zen , Florian Marquardt

A Markov decision problem is called reversible if the stationary controlled Markov chain is reversible under every stationary Markovian strategy. A natural application in which such problems arise is in the control of Metropolis-Hastings…

Probability · Mathematics 2022-07-13 Venkat Anantharam

Mean-Field Games are games with a continuum of players that incorporate the time-dimension through a control-theoretic approach. Recently, simpler approaches relying on the Best Reply Strategy have been proposed. They assume that the agents…

Optimization and Control · Mathematics 2014-12-24 Pierre Degond , Michael Herty , Jian-Guo Liu

This paper develops a policy gradient method for entropy-regularized mean-field control in the discounted infinite-horizon setting. We consider randomized feedback policies and a coupled representative-particle/population system, in which…

Optimization and Control · Mathematics 2026-05-21 Erhan Bayraktar , Martin Hernandez , Qinxin Yan , Yuhua Zhu

We investigate the existence of an optimal policy to monitor a mean field systems of agents managing a risky project under moral hazard with accidents modeled by L\'evy processes magnified by the law of the project. We provide a general…

Optimization and Control · Mathematics 2022-07-25 Thibaut Mastrolia , Jiacheng Zhang