English
Related papers

Related papers: Neural Actor-Critic Methods for Hamilton-Jacobi-Be…

200 papers

Policy iteration is a widely used technique to solve the Hamilton Jacobi Bellman (HJB) equation, which arises from nonlinear optimal feedback control theory. Its convergence analysis has attracted much attention in the unconstrained case.…

Optimization and Control · Mathematics 2020-05-19 Sudeep Kundu , Karl Kunisch

We study optimal control problems governed by abstract infinite dimensional stochastic differential equations using the dynamic programming approach. In the first part, we prove Lipschitz continuity, semiconcavity and semiconvexity of the…

Optimization and Control · Mathematics 2025-02-27 Filippo de Feo , Andrzej Święch , Lukas Wessels

The purpose of this note is to propose a new approach for the probabilistic interpretation of Hamilton-Jacobi-Bellman equations associated with stochastic recursive optimal control problems, utilizing the representation theorem for…

Probability · Mathematics 2017-05-03 Lishun Xiao , Shengjun Fan , Dejian Tian

Devising optimal interventions for diffusive systems often requires the solution of the Hamilton-Jacobi-Bellman (HJB) equation, a nonlinear backward partial differential equation (PDE), that is, in general, nontrivial to solve. Existing…

Statistical Mechanics · Physics 2022-10-18 Dimitra Maoutsa , Manfred Opper

We obtain weighted uniform estimates for the gradient of the solutions to a class of linear parabolic Cauchy problems with unbounded coefficients. Such estimates are then used to prove existence and uniqueness of the mild solution to a…

Analysis of PDEs · Mathematics 2014-02-04 Davide Addona

Optimal control of interacting particles governed by stochastic evolution equations in Hilbert spaces is an open area of research. Such systems naturally arise in formulations where each particle is modeled by stochastic partial…

Probability · Mathematics 2025-11-27 Filippo de Feo , Fausto Gozzi , Andrzej Święch , Lukas Wessels

In this work, we propose a class of numerical schemes for solving semilinear Hamilton-Jacobi-Bellman-Isaacs (HJBI) boundary value problems which arise naturally from exit time problems of diffusion processes with controlled drift. We…

Numerical Analysis · Mathematics 2020-02-14 Kazufumi Ito , Christoph Reisinger , Yufei Zhang

The objective of designing a control system is to steer a dynamical system with a control signal, guiding it to exhibit the desired behavior. The Hamilton-Jacobi-Bellman (HJB) partial differential equation offers a framework for optimal…

Machine Learning · Computer Science 2025-10-22 Jostein Barry-Straume , Adwait D. Verulkar , Arash Sarshar , Andrey A. Popov , Adrian Sandu

We consider a general class of stochastic optimal control problems, where the state process lives in a real separable Hilbert space and is driven by a cylindrical Brownian motion and a Poisson random measure; no special structure is imposed…

Probability · Mathematics 2018-10-04 Elena Bandini , Fulvia Confortola , Andrea Cosso

In this article, a class of optimal control problems of differential equations with delays are investigated for which the associated Hamilton-Jacobi-Bellman (HJB) equations are nonlinear partial differential equations with delays. This type…

Optimization and Control · Mathematics 2015-07-16 Jianjun Zhou

Hamilton-Jacobi (HJ) partial differential equations (PDEs) have diverse applications spanning physics, optimal control, game theory, and imaging sciences. This research introduces a first-order optimization-based technique for HJ PDEs,…

Numerical Analysis · Mathematics 2023-10-04 Tingwei Meng , Wenbo Hao , Siting Liu , Stanley J. Osher , Wuchen Li

We study a new two-time-scale stochastic gradient method for solving optimization problems, where the gradients are computed with the aid of an auxiliary variable under samples generated by time-varying MDPs controlled by the underlying…

Optimization and Control · Mathematics 2024-08-27 Sihan Zeng , Thinh T. Doan , Justin Romberg

We prove that a single-layer neural network trained with the online actor critic algorithm converges in distribution to a random ordinary differential equation (ODE) as the number of hidden units and the number of training steps…

Machine Learning · Computer Science 2026-05-28 Samuel Chun-Hei Lam , Justin Sirignano , Ziheng Wang

Many optimal control problems are formulated as two point boundary value problems (TPBVPs) with conditions of optimality derived from the Hamilton-Jacobi-Bellman (HJB) equations. In most cases, it is challenging to solve HJBs due to the…

Optimization and Control · Mathematics 2019-07-25 Sixiong You , Ran Dai , Ping Lu

Convex Q-learning is a recent approach to reinforcement learning, motivated by the possibility of a firmer theory for convergence, and the possibility of making use of greater a priori knowledge regarding policy or value function structure.…

Optimization and Control · Mathematics 2022-10-18 Fan Lu , Joel Mathias , Sean Meyn , Karanjit Kalsi

This paper deals with a class of neural SDEs and studies the limiting behavior of the associated sampled optimal control problems as the sample size grows to infinity. The neural SDEs with $N$ samples can be linked to the $N$-particle…

Optimization and Control · Mathematics 2025-06-19 Huafu Liao , Alpár R. Mészáros , Chenchen Mou , Chao Zhou

In this paper, we propose and study the stochastic path-dependent Hamilton-Jacobi-Bellman (SPHJB) equation that arises naturally from the optimal stochastic control problem of stochastic differential equations with path-dependence and…

Probability · Mathematics 2020-06-24 Jinniao Qiu

We present a simple and easy to implement method for the numerical solution of a rather general class of Hamilton-Jacobi-Bellman (HJB) equations. In many cases, the considered problems have only a viscosity solution, to which, fortunately,…

Computational Finance · Quantitative Finance 2011-02-17 Jan Hendrik Witte , Christoph Reisinger

By using an parametric value function to replace the Monte-Carlo rollouts for value estimation, the actor-critic (AC) algorithms can reduce the variance of stochastic policy gradient so that to improve the convergence rate. While existing…

Machine Learning · Computer Science 2024-08-19 Yanjie Dong , Haijun Zhang , Gang Wang , Shisheng Cui , Xiping Hu

In this article, we analyse optimal statistical arbitrage strategies from stochastic control and optimisation problems for multiple co-integrated stocks with eigenportfolios being factors. Optimal portfolio weights are found by solving a…

Portfolio Management · Quantitative Finance 2022-02-09 T. N. Li , A. Papanicolaou