English
Related papers

Related papers: Convergence of Policy Iteration for Entropy-Regula…

200 papers

This work considers the stability of nonlinear stochastic receding horizon control when the optimal controller is only computed approximately. A number of general classes of controller approximation error are analysed including…

Optimization and Control · Mathematics 2018-12-03 Francesco Bertoli , Adrian N. Bishop

Reinforcement Learning (RL) has emerged as a powerful framework for sequential decision-making in dynamic environments, particularly when system parameters are unknown. This paper investigates RL-based control for entropy-regularized…

Systems and Control · Electrical Eng. & Systems 2025-12-02 Gabriel Diaz , Lucky Li , Wenhao Zhang

We present a memory-bounded optimization approach for solving infinite-horizon decentralized POMDPs. Policies for each agent are represented by stochastic finite state controllers. We formulate the problem of optimizing these policies as a…

Artificial Intelligence · Computer Science 2012-06-26 Christopher Amato , Daniel S Bernstein , Shlomo Zilberstein

This paper focuses on the optimal control of a class of stochastic Volterra integral equations. Here the coefficients are regular and not assumed to be of convolution type. We show that, under mild regularity assumptions, these equations…

Probability · Mathematics 2026-04-08 Dylan Possamaï , Mehdi Talbi

The optimal \(H_{\infty}\) control problem over an infinite time horizon, which incorporates a performance function with a discount factor \(e^{-\alpha t}\) (\(\alpha > 0\)), is important in various fields. Solving this optimal…

Optimization and Control · Mathematics 2024-10-04 Guoyuan Chen , Yi Wang , Qinglong Zhou

We study an optimal control problem on infinite time horizon with semimartingale strategies, random coefficients and regime switching. The value function and the optimal strategy can be characterized in terms of three systems of backward…

Optimization and Control · Mathematics 2026-02-27 Xinman Cheng , Guanxing Fu , Xiaonyu Xia

In this note, we study a class of indefinite stochastic McKean-Vlasov linear-quadratic (LQ in short) control problem under the control taking nonnegative values. In contrast to the conventional issue, both the classical dynamic programming…

Optimization and Control · Mathematics 2023-10-05 Xun Li , Liangquan Zhang

A general time-inconsistent optimal control problem is considered for stochastic differential equations with deterministic coefficients. Under suitable conditions, a Hamilton-Jacobi-Bellman type equation is derived for the equilibrium value…

Optimization and Control · Mathematics 2012-04-04 Jiongmin Yong

The ergodic control problem for a non-degenerate controlled diffusion controlled through its drift is considered under a uniform stability condition that ensures the well-posedness of the associated Hamilton-Jacobi-Bellman (HJB) equation. A…

Optimization and Control · Mathematics 2019-03-20 Ari Arapostathis , Vivek S. Borkar

In this paper, we aim to solve the high dimensional stochastic optimal control problem from the view of the stochastic maximum principle via deep learning. By introducing the extended Hamiltonian system which is essentially an FBSDE with a…

Optimization and Control · Mathematics 2021-06-23 Shaolin Ji , Shige Peng , Ying Peng , Xichuan Zhang

In this manuscript, we study optimal control problems for stochastic delay differential equations using the dynamic programming approach in Hilbert spaces via viscosity solutions of the associated Hamilton-Jacobi-Bellman equations. We show…

Optimization and Control · Mathematics 2024-12-24 Filippo de Feo , Andrzej Święch

We study semi Lagrangian approximation schemes for Hamilton Jacobi Bellman equations arising from finite horizon optimal control problems. Classical error estimates for these schemes include the term $\frac{1}{\Delta t}$ which leads to…

Optimization and Control · Mathematics 2026-02-18 Alessandro Alla , Filippo Mayer

Adaptive optimal control using value iteration (VI) initiated from a stabilizing policy is theoretically analyzed in various aspects including the continuity of the result, the stability of the system operated using any single/constant…

Systems and Control · Computer Science 2015-05-18 Ali Heydari

We study optimal control problems governed by abstract infinite dimensional stochastic differential equations using the dynamic programming approach. In the first part, we prove Lipschitz continuity, semiconcavity and semiconvexity of the…

Optimization and Control · Mathematics 2025-02-27 Filippo de Feo , Andrzej Święch , Lukas Wessels

We consider a finite horizon stochastic optimal control problem for nearest-neighbor random walk $\{X_i\}$ on the set of integers. The cost function is the expectation of exponential of the path sum of a random stationary and ergodic…

Probability · Mathematics 2017-05-23 Atilla Yilmaz , Ofer Zeitouni

We establish central limit theorems for the Sample Average Approximation (SAA) method in discrete-time, finite-horizon stochastic optimal control. Our analysis is based on an abstract limit theorem for stochastic backward recursions, which…

Optimization and Control · Mathematics 2026-04-21 Johannes Milz , Alexander Shapiro

We consider the infinite-horizon discounted optimal control problem formalized by Markov Decision Processes. We focus on Policy Search algorithms, that compute an approximately optimal policy by following the standard Policy Iteration (PI)…

Artificial Intelligence · Computer Science 2013-06-04 Bruno Scherrer

Optimal control of stochastic nonlinear dynamical systems is a major challenge in the domain of robot learning. Given the intractability of the global control problem, state-of-the-art algorithms focus on approximate sequential optimization…

Machine Learning · Computer Science 2020-04-23 Joe Watson , Hany Abdulsamad , Jan Peters

Controlling the stochastic dynamics of biological populations is a challenge that arises across various biological contexts. However, these dynamics are inherently nonlinear and involve a discrete state space, i.e., the number of molecules,…

Populations and Evolution · Quantitative Biology 2025-10-21 Shuhei A. Horiguchi , Tetsuya J. Kobayashi

We propose a scalable, policy-centric framework for continuous-time multi-asset portfolio-consumption optimization under inequality constraints. Our method integrates neural policies with Pontryagin's Maximum Principle (PMP) and enforces…

Portfolio Management · Quantitative Finance 2025-11-07 Jeonggyu Huh , Jaegi Jeon , Hyeng Keun Koo , Byung Hwa Lim