English
Related papers

Related papers: On Dynamic Programming Theory for Leader-Follower …

200 papers

We investigate the linear quadratic Gaussian Stackelberg game under a class of nested observation information pattern. Two decision makers implement control strategies relying on different information sets: The follower uses its observation…

Optimization and Control · Mathematics 2022-06-07 Zhipeng Li , Damian Marelli , Minyue Fu , Huanshui Zhang

We consider a dynamic programming (DP) approach to approximately solving an infinite-horizon constrained Markov decision process (CMDP) problem with a fixed initial-state for the expected total discounted-reward criterion with a…

Optimization and Control · Mathematics 2023-08-08 Hyeong Soo Chang

Within the framework of viscosity solution, we study the relationship between the maximum principle (MP) in [9] and the dynamic programming principle (DPP) in [10] for a fully coupled forward-backward stochastic controlled system (FBSCS)…

Optimization and Control · Mathematics 2018-05-17 Mingshang Hu , Shaolin Ji , Xiaole Xue

Markov decision processes (MDP) are a well-established model for sequential decision-making in the presence of probabilities. In robust MDP (RMDP), every action is associated with an uncertainty set of probability distributions, modelling…

Artificial Intelligence · Computer Science 2024-12-16 Tobias Meggendorfer , Maximilian Weininger , Patrick Wienhöft

This paper is concerned with a Stackelberg stochastic differential game, where the systems are driven by stochastic differential equation (SDE for short), in which the control enters the randomly disturbed coefficients (drift and…

Optimization and Control · Mathematics 2021-08-12 Liangquan Zhang , Wei Zhang

The multi-leader--multi-follower game (MLMFG) involves two or more leaders and followers and serves as a generalization of the Stackelberg game and the single-leader--multi-follower game (SLMFG). Although MLMFG covers wide range of…

Optimization and Control · Mathematics 2024-04-09 Atsushi Hori , Daisuke Tsuyuguchi , Ellen H. Fukuda

We study a two-player discounted zero-sum stochastic game model for dynamic operational planning in military campaigns. At each stage, the players manage multiple commanders who order military actions on objectives that have an open line of…

Computer Science and Game Theory · Computer Science 2024-03-04 Joseph E. McCarthy , Mathieu Dahan , Chelsea C. White

This paper presents a new theory, known as robust dynamic pro- gramming, for a class of continuous-time dynamical systems. Different from traditional dynamic programming (DP) methods, this new theory serves as a fundamental tool to analyze…

Optimization and Control · Mathematics 2018-09-18 Tao Bian , Zhong-Ping Jiang

This paper is concerned with the stochastic linear quadratic Stackelberg differential game with overlapping information, where the diffusion terms contain the control and state variables. Here the term "overlapping" means that there are…

Optimization and Control · Mathematics 2018-05-01 Jingtao Shi , Guangchen Wang , Jie Xiong

A recent theory shows that a multi-player decentralized partially observable Markov decision process can be transformed into an equivalent single-player game, enabling the application of \citeauthor{bellman}'s principle of optimality to…

Computer Science and Game Theory · Computer Science 2025-01-03 Johan Peralez , Aurélien Delage , Olivier Buffet , Jilles S. Dibangoye

In this paper, a leader-follower stochastic differential game is studied for a linear stochastic differential equation with a quadratic cost functional. The coefficients in the state equation and the weighting matrices in the cost…

Optimization and Control · Mathematics 2021-07-13 Zixuan Li , Jingtao Shi

In many practical uses of reinforcement learning (RL) the set of actions available at a given state is a random variable, with realizations governed by an exogenous stochastic process. Somewhat surprisingly, the foundations for such…

Artificial Intelligence · Computer Science 2021-02-16 Craig Boutilier , Alon Cohen , Amit Daniely , Avinatan Hassidim , Yishay Mansour , Ofer Meshi , Martin Mladenov , Dale Schuurmans

This paper is concerned with a linear-quadratic partially observed Stackelberg stochastic differential game with correlated state and observation noises, where the diffusion coefficient does not contain the control variable and the control…

Optimization and Control · Mathematics 2021-05-25 Yueyang Zheng , Jingtao Shi

We study Recursive Concurrent Stochastic Games (RCSGs), extending our recent analysis of recursive simple stochastic games to a concurrent setting where the two players choose moves simultaneously and independently at each state. For…

Computer Science and Game Theory · Computer Science 2015-07-01 Kousha Etessami , Mihalis Yannakakis

We consider the repeated prisoner's dilemma (PD). We assume that players make their choices knowing only average payoffs from the previous stages. A player's strategy is a function from the convex hull $\mathfrak{S}$ of the set of payoffs…

Optimization and Control · Mathematics 2018-05-16 Sławomir Plaskacz , Joanna Zwierzchowska

This paper introduces a class of continuous-time, finite-player stochastic general-sum differential games that admit solutions through an exact linear PDE system. We formulate a distribution planning game utilizing the cross-log-likelihood…

Optimization and Control · Mathematics 2026-04-10 Monika Tomar , Takashi Tanaka

Dynamic programming (DP) is a fundamental tool used across many engineering fields. The main goal of DP is to solve Bellman's optimality equations for a given Markov decision process (MDP). Standard methods like policy iteration exploit the…

Artificial Intelligence · Computer Science 2025-07-30 Sergio Rozada , Samuel Rey , Gonzalo Mateos , Antonio G. Marques

We study a class of stochastic target games where one player tries to find a strategy such that the state process almost-surely reaches a given target, no matter which action is chosen by the opponent. Our main result is a geometric dynamic…

Probability · Mathematics 2015-02-03 Bruno Bouchard , Marcel Nutz

Markov Decision Processes (MDPs) are a formal framework for modeling and solving sequential decision-making problems. In finite-time horizons such problems are relevant for instance for optimal stopping or specific supply chain problems,…

Optimization and Control · Mathematics 2024-05-07 Sara Klein , Simon Weissmann , Leif Döring

The concept of leader--follower (or Stackelberg) equilibrium plays a central role in a number of real--world applications of game theory. While the case with a single follower has been thoroughly investigated, results with multiple…

Computer Science and Game Theory · Computer Science 2017-07-10 Nicola Basilico , Stefano Coniglio , Nicola Gatti