English
Related papers

Related papers: Neural Actor-Critic Methods for Hamilton-Jacobi-Be…

200 papers

We develop a new policy gradient and actor-critic algorithm for solving mean-field control problems within a continuous time reinforcement learning setting. Our approach leverages a gradient-based representation of the value function,…

Machine Learning · Statistics 2023-09-11 Huyên Pham , Xavier Warin

Solving high dimensional optimal control problems and corresponding Hamilton-Jacobi PDEs are important but challenging problems in control engineering. In this paper, we propose two abstract neural network architectures which are…

Optimization and Control · Mathematics 2023-03-31 Jérôme Darbon , Peter M. Dower , Tingwei Meng

A gradient-enhanced functional tensor train cross approximation method for the resolution of the Hamilton-Jacobi-Bellman (HJB) equations associated to optimal feedback control of nonlinear dynamics is presented. The procedure uses samples…

Numerical Analysis · Mathematics 2023-02-23 Sergey Dolgov , Dante Kalise , Luca Saluzzi

We study the problem of generating control laws for systems with unknown dynamics. Our approach is to represent the controller and the value function with neural networks, and to train them using loss functions adapted from the…

Robotics · Computer Science 2023-02-21 Selim Engin , Volkan Isler

The Bellman equation and its continuous form, the Hamilton-Jacobi-Bellman equation, are ubiquitous in reinforcement learning and control theory. However, these equations become intractable for high-dimensional or nonlinear systems. This…

Artificial Intelligence · Computer Science 2026-05-04 Preston Rozwood , Edward Mehrez , Ludger Paehler , Wen Sun , Steven L. Brunton

We develop the dynamic programming approach for a family of infinite horizon boundary control problems with linear state equation and convex cost. We prove that the value function of the problem is the unique regular solution of the…

Optimization and Control · Mathematics 2008-06-27 Silvia Faggian , Fausto Gozzi

We propose a machine learning algorithm for solving finite-horizon stochastic control problems based on a deep neural network representation of the optimal policy functions. The algorithm has three features: (1) It can solve…

General Economics · Economics 2024-12-09 Xianhua Peng , Steven Kou , Lekang Zhang

This paper addresses the numerical solution of backward stochastic differential equations (BSDEs) arising in stochastic optimal control. Specifically, we investigate two BSDEs: one derived from the Hamilton-Jacobi-Bellman equation and the…

Optimization and Control · Mathematics 2025-03-12 Yuhang Mei , Amirhossein Taghvaei

In this paper, we study a stochastic recursive optimal control problem in which the value functional is defined by the solution of a backward stochastic differential equation (BSDE) under $\tilde{G}$-expectation. Under standard assumptions,…

Optimization and Control · Mathematics 2021-06-08 Mingshang Hu , Shaolin Ji , Xiaojuan Li

In this paper, we study one kind of stochastic recursive optimal control problem with the obstacle constraints for the cost function where the cost function is described by the solution of one reflected backward stochastic differential…

Optimization and Control · Mathematics 2007-05-23 Zhen Wu , Zhiyong Yu

In this paper, we mainly focus on solving high-dimensional stochastic Hamiltonian systems with boundary condition, which is essentially a Forward Backward Stochastic Differential Equation (FBSDE in short), and propose a novel method from…

Optimization and Control · Mathematics 2021-12-13 Shaolin Ji , Shige Peng , Ying Peng , Xichuan Zhang

We study a stochastic optimal control problem with the state constrained to a smooth, compact domain. The control influences both the drift and a possibly degenerate, control-dependent dispersion matrix, leading to a fully nonlinear,…

Optimization and Control · Mathematics 2025-08-08 Anderson O. Calixto , Bernardo Freitas Paulo da Costa , Glauco Valle

In this paper, we study a stochastic recursive optimal control problem in which the system is governed by a functional forward-backward stochastic differential equation. Under standard assumptions, we establish the dynamic programming…

Probability · Mathematics 2013-01-03 Shaolin Ji , Shuzhen Yang

The solution to a stochastic optimal control problem can be determined by computing the value function from a discretization of the associated Hamilton-Jacobi-Bellman equation. Alternatively, the problem can be reformulated in terms of a…

Optimization and Control · Mathematics 2024-02-29 Sebastian Reich

We address the discounted reward setting in reinforcement learning (RL). To mitigate the value approximation challenges in policy gradient methods, actor-critic approaches have been developed and are known to converge to stationary points…

Machine Learning · Computer Science 2026-05-15 Sanjeev Manivannan , Shuban V

Choosing how much noise to add in Langevin dynamics is essential for making these algorithms effective in challenging optimization problems. One promising approach is to determine this noise by solving Hamilton-Jacobi-Bellman (HJB)…

Numerical Analysis · Mathematics 2026-03-19 Taorui Wang , Xun Li , Gu Wang , Zhongqiang Zhang

Policy iteration (PI) is a widely used algorithm for synthesizing optimal feedback control policies across many engineering and scientific applications. When PI is deployed on infinite-horizon, nonlinear, autonomous optimal-control…

Optimization and Control · Mathematics 2025-07-15 Tobias Ehring , Behzad Azmi , Bernard Haasdonk

In this manuscript we consider optimal control problems of stochastic differential equations with delays in the state and in the control. First, we prove an equivalent Markovian reformulation on Hilbert spaces of the state equation. Then,…

Optimization and Control · Mathematics 2024-05-20 Filippo de Feo

This paper proposes a new actor-critic-style algorithm called Dual Actor-Critic or Dual-AC. It is derived in a principled way from the Lagrangian dual form of the Bellman optimality equation, which can be viewed as a two-player game between…

Machine Learning · Computer Science 2018-01-01 Bo Dai , Albert Shaw , Niao He , Lihong Li , Le Song

We study optimal stochastic control problems of general coupled systems of forward-backward stochastic differential equations with jumps. By means of the It\^o-Ventzell formula the system is transformed to a controlled backward stochastic…

Optimization and Control · Mathematics 2017-01-12 Bernt Øksendal , Agnès Sulem , Tusheng Zhang
‹ Prev 1 8 9 10 Next ›