中文
相关论文

相关论文: Non-Markovian Impulse Control Under Nonlinear Expe…

200 篇论文

We study a nonzero-sum stochastic differential game with both players adopting impulse controls, on a finite time horizon. The objective of each player is to maximize her total expected discounted profits. The resolution methodology relies…

最优化与控制 · 数学 2021-12-21 René Aïd , Lamia Ben Ajmia , M'hamed Gaïgi , Mohamed Mnif

This paper aims to study the relationship between the maximum principle and the dynamic programming principle for recursive optimal control problem of stochastic evolution equations, where the control domain is not necessarily convex and…

最优化与控制 · 数学 2025-12-19 Ying Hu , Guomin Liu , Shanjian Tang

This paper considers an optimal impulse control problem of dynamical systems generated by a flow. The performance criteria are total costs over the infinite time horizon. Apart from the main performance to be minimized, there are multiple…

最优化与控制 · 数学 2020-10-27 Alexey Piunovskiy , Yi Zhang

This paper studies a class of non$-$Markovian singular stochastic control problems, for which we provide a novel probabilistic representation. The solution of such control problem is proved to identify with the solution of a $Z-$constrained…

最优化与控制 · 数学 2018-02-27 Romuald Elie , Ludovic Moreau , Dylan Possamaï

We consider two-player zero-sum differential games (ZSDGs), where the state process (dynamical system) depends on the random initial condition and the state process's distribution, and the objective functional includes the state process's…

最优化与控制 · 数学 2020-05-26 Jun Moon , Tamer Basar

In this article, we discuss two algorithms tailored to discrete-time deterministic finite-horizon nonlinear optimal control problems or so-called deterministic trajectory optimization problems. Both algorithms can be derived from an…

最优化与控制 · 数学 2024-12-10 Mohammad Mahmoudi Filabadi , Tom Lefebvre , Guillaume Crevecoeur

We present differentiable predictive control (DPC), a method for learning constrained neural control policies for linear systems with probabilistic performance guarantees. We employ automatic differentiation to obtain direct policy…

系统与控制 · 电气工程与系统科学 2022-01-28 Jan Drgona , Aaron Tuor , Draguna Vrabie

We study the optimal control of path-dependent McKean-Vlasov equations valued in Hilbert spaces motivated by non Markovian mean-field models driven by stochastic PDEs. We first establish the well-posedness of the state equation, and then we…

最优化与控制 · 数学 2022-12-21 Andrea Cosso , Fausto Gozzi , Idris Kharroubi , Huyên Pham , Mauro Rosestolato

We study a class of zero-sum games between a singular-controller and a stopper over finite-time horizon. The underlying process is a multi-dimensional (locally non-degenerate) controlled stochastic differential equation (SDE) evolving in an…

最优化与控制 · 数学 2023-10-31 Andrea Bovo , Tiziano De Angelis , Elena Issoglio

Dynamic games arise when multiple agents with differing objectives control a dynamic system. They model a wide variety of applications in economics, defense, energy systems and etc. However, compared to single-agent control problems, the…

系统与控制 · 电气工程与系统科学 2020-01-08 Bolei Di , Andrew Lamperski

We study optimal stochastic control problem for non-Markovian stochastic differential equations (SDEs) where the drift, diffusion coefficients, and gain functionals are path-dependent, and importantly we do not make any ellipticity…

概率论 · 数学 2013-11-04 Marco Fuhrman , Huyên Pham

Markov decision processes (MDPs) are a popular model for performance analysis and optimization of stochastic systems. The parameters of stochastic behavior of MDPs are estimates from empirical observations of a system; their values are not…

人工智能 · 计算机科学 2017-10-26 Dimitri Scheftelowitsch , Peter Buchholz , Vahid Hashemi , Holger Hermanns

Within the framework of viscosity solution, we study the relationship between the maximum principle (MP) in [9] and the dynamic programming principle (DPP) in [10] for a fully coupled forward-backward stochastic controlled system (FBSCS)…

最优化与控制 · 数学 2018-05-17 Mingshang Hu , Shaolin Ji , Xiaole Xue

In this paper, we study a class of zero-sum two-player stochastic differential games with the controlled stochastic differential equations and the payoff/cost functionals of recursive type. As opposed to the pioneering work by Fleming and…

概率论 · 数学 2021-05-21 Jinniao Qiu , Jing Zhang

In dynamic noncooperative games, each player makes conjectures about other players' reactions before choosing a strategy. However, resulting equilibria may be multiple and do not always lead to desirable outcomes. These issues are typically…

计算机科学与博弈论 · 计算机科学 2025-11-24 Francesco Morri , Hélène Le Cadre , David Salas , Didier Aussel

Differential Dynamic Programming (DDP) is an efficient computational tool for solving nonlinear optimal control problems. It was originally designed as a single shooting method and thus is sensitive to the initial guess supplied. This work…

机器人学 · 计算机科学 2023-09-29 He Li , Wenhao Yu , Tingnan Zhang , Patrick M. Wensing

In this paper we study zero-sum two-player stochastic differential games with jumps with the help of theory of Backward Stochastic Differential Equations (BSDEs). We generalize the results of Fleming and Souganidis [10] and those by Biswas…

最优化与控制 · 数学 2010-04-19 Rainer Buckdahn , Ying Hu , Juan Li

In several standard models of dynamic programming (gambling houses, MDPs, POMDPs), we prove the existence of a very robust notion of value for the infinitely repeated problem, namely the pathwise uniform value. This solves two open…

最优化与控制 · 数学 2015-09-09 Xavier Venel , Bruno Ziliotto

This paper investigates the relationship between Pontryagin's maximum principle and dynamic programming principle in the context of stochastic optimal control systems governed by stochastic evolution equations with random coefficients in…

最优化与控制 · 数学 2025-11-05 Dingqian Gao , Qi Lü

This work presents DMPC (Data-and Model-Driven Predictive Control) to solve control problems in which some of the constraints or parts of the objective function are known, while others are entirely unknown to the controller. It is assumed…

系统与控制 · 电气工程与系统科学 2021-03-02 Hassan Jafarzadeh , Cody Fleming