English
Related papers

Related papers: Non-Markovian Impulse Control Under Nonlinear Expe…

200 papers

Leader-follower general-sum stochastic games (LF-GSSGs) model sequential decision-making under asymmetric commitment, where a leader commits to a policy and a follower best responds, yielding a strong Stackelberg equilibrium (SSE) with…

Computer Science and Game Theory · Computer Science 2025-12-08 Jilles Steeve Dibangoye , Thibaut Le Marre , Ocan Sankur , François Schwarzentruber

In this paper, we study a stochastic optimal control problem under degenerate G-expectation. By using implied partition method, we show that the approximation result for admissible controls still hold. Based on this result, we prove that…

Optimization and Control · Mathematics 2022-10-19 Xiaojuan Li

We show that the value function of a stochastic control problem is the unique solution of the associated Hamilton-Jacobi-Bellman (HJB) equation, completely avoiding the proof of the so-called dynamic programming principle (DPP). Using…

Probability · Mathematics 2013-09-25 Erhan Bayraktar , Mihai Sirbu

In this work, we study dynamic programming (DP) algorithms for partially observable Markov decision processes with jointly continuous and discrete state-spaces. We consider a class of stochastic systems which have coupled discrete and…

Optimization and Control · Mathematics 2019-03-07 Donghwan Lee , Niao He , Jianghai Hu

We study a robust optimal stopping problem with respect to a set $\cP$ of mutually singular probabilities. This can be interpreted as a zero-sum controller-stopper game in which the stopper is trying to maximize its pay-off while an adverse…

Probability · Mathematics 2016-04-12 Erhan Bayraktar , Song Yao

Mirror play (MP) is a well-accepted primal-dual multi-agent learning algorithm where all agents simultaneously implement mirror descent in a distributed fashion. The advantage of MP over vanilla gradient play lies in its usage of mirror…

Computer Science and Game Theory · Computer Science 2024-03-26 Yunian Pan , Tao Li , Quanyan Zhu

When a vehicle drives on the road, its behaviors will be affected by surrounding vehicles. Prediction and decision should not be considered as two separate stages because all vehicles make decisions interactively. This paper constructs the…

Artificial Intelligence · Computer Science 2023-02-09 Xujie Song , Zexi Lin

Safety in stochastic control systems, which are subject to random noise with a known probability distribution, aims to compute policies that satisfy predefined operational constraints with high confidence throughout the uncertain evolution…

Systems and Control · Electrical Eng. & Systems 2025-11-12 Saber Omidi , Marek Petrik , Se Young Yoon , Momotaz Begum

Several attempts to dampen the curse of dimensionnality problem of the Dynamic Programming approach for solving multistage optimization problems have been investigated. One popular way to address this issue is the Stochastic Dual Dynamic…

Optimization and Control · Mathematics 2020-10-09 Marianne Akian , Jean-Philippe Chancelier , Benoît Tran

We propose empirical dynamic programming algorithms for Markov decision processes (MDPs). In these algorithms, the exact expectation in the Bellman operator in classical value iteration is replaced by an empirical estimate to get `empirical…

Optimization and Control · Mathematics 2013-11-26 William B. Haskell , Rahul Jain , Dileep Kalathil

We consider the inverse problem of dynamic games, where cost function parameters are sought which explain observed behavior of interacting players. Maximum entropy inverse reinforcement learning is extended to the N-player case in order to…

Systems and Control · Electrical Eng. & Systems 2020-07-27 Jairo Inga , Esther Bischoff , Florian Köpf , Sören Hohmann

Since Peng (1993) established a local maximum principle for a general stochastic control problem governed by forward-backward stochastic differential equations (FBSDEs), the corresponding partial differential equation (PDE) characterization…

Optimization and Control · Mathematics 2025-08-07 Yuhong Xu , Shuzhen Yang

Dynamic Mode Decomposition (DMD) has emerged as a powerful tool for analyzing the dynamics of non-linear systems from experimental datasets. Recently, several attempts have extended DMD to the context of low-rank approximations. This…

Machine Learning · Statistics 2018-05-18 Patrick Héas , Cédric Herzet

We study zero-sum stochastic differential games with player dynamics governed by a nondegenerate controlled diffusion process. Under the assumption of uniform stability, we establish the existence of a solution to the Isaac's equation for…

Optimization and Control · Mathematics 2019-03-20 Ari Arapostathis , Vivek S. Borkar , K. Suresh Kumar

The paper extends an impulsive control-theoretical framework towards dynamic systems in the space of measures. We consider a transport equation describing the time-evolution of a conservative "mass" (probability measure), which represents…

Optimization and Control · Mathematics 2020-02-17 Nikolay Pogodaev , Maxim Staritsyn

We consider two classes of constrained finite state-action stochastic games. First, we consider a two player nonzero sum single controller constrained stochastic game with both average and discounted cost criterion. We consider the same…

Optimization and Control · Mathematics 2012-06-11 Vikas Vikram Singh , N. Hemachandra

We consider a multi-period stochastic control problem where the multivariate driving stochastic factor of the system has known marginal distributions but uncertain dependence structure. To solve the problem, we propose to implement the…

Optimization and Control · Mathematics 2022-09-13 Erhan Bayraktar , Tao Chen

This investigation is dedicated to a two-player zero-sum stochastic differential game (SDG), where a cost function is characterized by a backward stochastic differential equation (BSDE) with a continuous and monotonic generator regarding…

Optimization and Control · Mathematics 2024-04-19 Guangchen Wang , Zhuangzhuang Xing

This paper introduces a novel Differential Dynamic Programming (DDP) algorithm for solving discrete-time finite-horizon optimal control problems with inequality constraints. Two variants, namely Feasible- and Infeasible-IPDDP algorithms,…

Systems and Control · Electrical Eng. & Systems 2020-10-21 Andrei Pavlov , Iman Shames , Chris Manzie

We study the optimal control of general stochastic McKean-Vlasov equation. Such problem is motivated originally from the asymptotic formulation of cooperative equilibrium for a large population of particles (players) in mean-field…

Probability · Mathematics 2017-01-06 Huyên Pham , Xiaoli Wei
‹ Prev 1 8 9 10 Next ›