English
Related papers

Related papers: A neural network based policy iteration algorithm …

200 papers

Physics-informed neural solvers offer a promising route to model-based reinforcement learning in continuous time, where optimal feedback synthesis is governed by Hamilton--Jacobi--Bellman (HJB) equations. Practical implementations often…

Machine Learning · Computer Science 2026-05-11 Minseok Kim , Yeongjong Kim , Namkyeong Cho , Yeoneung Kim

We analyse two practical aspects that arise in the numerical solution of Hamilton-Jacobi-Bellman (HJB) equations by a particular class of monotone approximation schemes known as semi-Lagrangian schemes. These schemes make use of a wide…

Numerical Analysis · Mathematics 2016-11-08 Christoph Reisinger , Julen Rotaetxe Arto

The approximation of solutions to second order Hamilton--Jacobi--Bellman (HJB) equations by deep neural networks is investigated. It is shown that for HJB equations that arise in the context of the optimal control of certain Markov…

Numerical Analysis · Mathematics 2021-03-11 Philipp Grohs , Lukas Herrmann

We propose a machine learning algorithm for solving finite-horizon stochastic control problems based on a deep neural network representation of the optimal policy functions. The algorithm has three features: (1) It can solve…

General Economics · Economics 2024-12-09 Xianhua Peng , Steven Kou , Lekang Zhang

Direct numerical simulation of diffusion through heterogeneous media can be difficult due to the computational cost of resolving fine-scale heterogeneities. One method to overcome this difficulty is to homogenize the model by replacing the…

Numerical Analysis · Mathematics 2020-11-24 Nathan G. March , Elliot J. Carr , Ian W. Turner

A learning technique for finite horizon optimal control problems and its approximation based on polynomials is analyzed. It allows to circumvent, in part, the curse dimensionality which is involved when the feedback law is constructed by…

Optimization and Control · Mathematics 2023-02-21 Karl Kunisch , Donato Vásquez-Varas

In this paper we study the fully nonlinear stochastic Hamilton-Jacobi-Bellman (HJB) equation for the optimal stochastic control problem of stochastic differential equations with random coefficients. The notion of viscosity solution is…

Optimization and Control · Mathematics 2018-07-16 Jinniao Qiu

We propose a neural network approach that yields approximate solutions for high-dimensional optimal control problems and demonstrate its effectiveness using examples from multi-agent path finding. Our approach yields controls in a feedback…

Optimization and Control · Mathematics 2022-06-29 Derek Onken , Levon Nurbekyan , Xingjian Li , Samy Wu Fung , Stanley Osher , Lars Ruthotto

We study continuous-time heterogeneous agent models cast as Mean Field Games, in the Aiyagari-Bewley-Huggett framework. The model couples a Hamilton-Jacobi-Bellman equation for individual optimization with a Fokker-Planck-Kolmogorov…

Optimization and Control · Mathematics 2025-10-02 Fabio Camilli , Qing Tang , Yong-shen Zhou

In this paper, we study a time-inconsistent stochastic optimal control problem with a recursive cost functional by a multi-person hierarchical differential game approach. An equilibrium strategy of this problem is constructed and a…

Optimization and Control · Mathematics 2016-06-13 Qingmeng Wei , Jiongmin Yong , Zhiyong Yu

This paper presents a physics-informed machine learning approach for synthesizing optimal feedback control policy for infinite-horizon optimal control problems by solving the Hamilton-Jacobi-Bellman (HJB) partial differential equation(PDE).…

Systems and Control · Electrical Eng. & Systems 2025-11-24 Tanay Raghunandan Srinivasa , Suraj Kumar

We introduce a new numerical method to approximate the solution of a finite horizon deterministic optimal control problem. We exploit two Hamilton-Jacobi-Bellman PDE, arising by considering the dynamics in forward and backward time. This…

Optimization and Control · Mathematics 2023-04-21 Marianne Akian , Stéphane Gaubert , Shanqing Liu

We investigate the long time behavior of weakly dissipative semilinear Hamilton-Jacobi-Bellman (HJB) equations and the turnpike property for the corresponding stochastic control problems. To this aim, we develop a probabilistic approach…

Probability · Mathematics 2023-03-17 Giovanni Conforti

This paper presents a new methodology to craft navigation functions for nonlinear systems with stochastic uncertainty. The method relies on the transformation of the Hamilton-Jacobi-Bellman (HJB) equation into a linear partial differential…

Robotics · Computer Science 2014-09-23 Matanya B. Horowitz , Joel W. Burdick

An advantageous feature of piecewise constant policy timestepping for Hamilton-Jacobi-Bellman (HJB) equations is that different linear approximation schemes, and indeed different meshes, can be used for the resulting linear equations for…

Numerical Analysis · Mathematics 2016-01-21 Christoph Reisinger , Peter Forsyth

This paper develops algorithms for high-dimensional stochastic control problems based on deep learning and dynamic programming. Unlike classical approximate dynamic programming approaches, we first approximate the optimal policy by means of…

Probability · Mathematics 2021-09-21 Côme Huré , Huyên Pham , Achref Bachouch , Nicolas Langrené

Continuous-time stochastic processes underlie many natural and engineered systems. In healthcare, autonomous driving, and industrial control, direct interaction with the environment is often unsafe or impractical, motivating offline…

Machine Learning · Statistics 2025-11-14 Nicolas Hoischen , Petar Bevanda , Max Beier , Stefan Sosnowski , Boris Houska , Sandra Hirche

We study the approximation of parabolic Hamilton-Jacobi-Bellman (HJB) equations in bounded domains with strong Dirichlet boundary conditions. We work under the assumption of the existence of a sufficiently regular barrier function for the…

Numerical Analysis · Mathematics 2019-07-16 Athena Picarelli , Christoph Reisinger , Julen Rotaetxe Arto

We study the problem of optimal portfolio selection under stochastic volatility within a continuous time reinforcement learning framework with portfolio constraints. Exploration is modeled through entropy-regularized relaxed controls, where…

Mathematical Finance · Quantitative Finance 2026-04-27 Thai Nguyen , Pertiny Nkuize

The solution to a stochastic optimal control problem can be determined by computing the value function from a discretization of the associated Hamilton-Jacobi-Bellman equation. Alternatively, the problem can be reformulated in terms of a…

Optimization and Control · Mathematics 2024-02-29 Sebastian Reich