English
Related papers

Related papers: A neural network based policy iteration algorithm …

200 papers

We prove a stochastic representation formula for the viscosity solution of Dirichlet terminal-boundary value problem for a degenerate Hamilton-Jacobi-Bellman integro-partial differential equation in a bounded domain. We show that the unique…

Probability · Mathematics 2018-08-23 Ruoting Gong , Chenchen Mou , Andrzej Swiech

The objective of designing a control system is to steer a dynamical system with a control signal, guiding it to exhibit the desired behavior. The Hamilton-Jacobi-Bellman (HJB) partial differential equation offers a framework for optimal…

Machine Learning · Computer Science 2025-10-22 Jostein Barry-Straume , Adwait D. Verulkar , Arash Sarshar , Andrey A. Popov , Adrian Sandu

In this paper, we propose a martingale-based neural network, SOC-MartNet, for solving high-dimensional Hamilton-Jacobi-Bellman (HJB) equations where no explicit expression is needed for the infimum of the Hamiltonian, $\inf_{u \in U}…

Numerical Analysis · Mathematics 2025-03-18 Wei Cai , Shuixin Fang , Tao Zhou

We investigate an optimal control problem for a diffusion whose drift and running cost are merely measurable in the state variable. Such low regularity rules out the use of Pontryagin's maximum principle and also invalidates the standard…

Optimization and Control · Mathematics 2025-09-03 Kai Du , Qingmeng Wei

We develop the dynamic programming approach for a family of infinite horizon boundary control problems with linear state equation and convex cost. We prove that the value function of the problem is the unique regular solution of the…

Optimization and Control · Mathematics 2008-06-27 Silvia Faggian , Fausto Gozzi

This work presents a novel policy iteration algorithm to tackle nonzero-sum stochastic impulse games arising naturally in many applications. Despite the obvious impact of solving such problems, there are no suitable numerical methods…

Optimization and Control · Mathematics 2020-06-29 René Aïd , Francisco Bernal , Mohamed Mnif , Diego Zabaljauregui , Jorge P. Zubelli

This paper investigates the convergence properties of the upwind difference scheme for the Hamilton--Jacobi--Bellman (HJB) equation, a central partial differential equation in optimal control theory. First, assuming the existence of a…

Numerical Analysis · Mathematics 2026-02-05 Daisuke Inoue , Yuji Ito , Takahito Kashiwabara , Norikazu Saito , Hiroaki Yoshida

The $H_\infty$ control design problem is considered for nonlinear systems with unknown internal system model. It is known that the nonlinear $ H_\infty $ control problem can be transformed into solving the so-called Hamilton-Jacobi-Isaacs…

Systems and Control · Computer Science 2014-05-13 Biao Luo , Huai-Ning Wu , Tingwen Huang

Optimal feedback controllers for nonlinear systems can be derived by solving the Hamilton-Jacobi-Bellman (HJB) equation. However, because the HJB is a nonlinear partial differential equation, numerical methods typically provide only…

Optimization and Control · Mathematics 2026-03-25 Morgan Jones , Matthew Peet

Mixed optimal stopping and stochastic control problems define variational inequalities with non-linear Hamilton-Jacobi-Bellman (HJB) operators, whose numerical solution is notoriously difficult and lack of reliable benchmarks. We first use…

Optimization and Control · Mathematics 2025-05-27 Yun Zhao , Harry Zheng

This paper presents SIMPOL (Simplified Policy Iteration), a modular numerical framework for solving continuous-time heterogeneous agent models. The core economic problem, the optimization of consumption and savings under idiosyncratic…

Computational Finance · Quantitative Finance 2025-09-30 Ricardo Alonzo Fernández Salguero

From the Hamilton-Jacobi-Bellman equation for the value function we derive a non-linear partial differential equation for the optimal portfolio strategy (the dynamic control). The equation is general in the sense that it does not depend on…

Portfolio Management · Quantitative Finance 2013-11-20 Mads Nielsen

We address finding the semi-global solutions to optimal feedback control and the Hamilton--Jacobi--Bellman (HJB) equation. Using the solution of an HJB equation, a feedback optimal control law can be implemented in real-time with minimum…

Optimization and Control · Mathematics 2016-06-17 Wei Kang , Lucas C. Wilcox

In this paper we discuss energy conservation issues related to the numerical solution of the nonlinear wave equation. As is well known, this problem can be cast as a Hamiltonian system that may be autonomous or not, depending on the…

Numerical Analysis · Mathematics 2017-11-27 Luigi Brugnano , Gianluca Frasca Caccia , Felice Iavernaro

We introduce a generic numerical schemes for fully nonlinear parabolic PDEs on the full domain, where the nonlinearity is convex on the Hessian of the solution. The main idea behind this paper is reduction of a fully nonlinear problem to a…

Analysis of PDEs · Mathematics 2024-10-08 Hung Duong , Arash Fahim

This paper introduces a reinforcement learning-based tracking control approach for a class of nonlinear systems using neural networks. In this approach, adversarial attacks were considered both in the actuator and on the outputs. This…

Systems and Control · Electrical Eng. & Systems 2022-09-20 Farshad Rahimi , Sepideh Ziaei

This paper considers optimal control of dynamical systems which are represented by nonlinear stochastic differential equations. It is well-known that the optimal control policy for this problem can be obtained as a function of a value…

Robotics · Computer Science 2014-05-30 Oktay Arslan , Evangelos Theodorou , Panagiotis Tsiotras

We establish a connection between stochastic optimal control and generative models based on stochastic differential equations (SDEs), such as recently developed diffusion probabilistic models. In particular, we derive a…

Machine Learning · Computer Science 2024-03-27 Julius Berner , Lorenz Richter , Karen Ullrich

We generalize the technique of [Solving Dirichlet boundary-value problems on curved domains by extensions from subdomains, SIAM J. Sci. Comput. 34, pp. A497--A519 (2012)] to elliptic problems with mixed boundary conditions and elliptic…

Numerical Analysis · Mathematics 2015-11-24 Weifeng Qiu , Manuel Solano , Patrick Vega

We present a midpoint policy iteration algorithm to solve linear quadratic optimal control problems in both model-based and model-free settings. The algorithm is a variation of Newton's method, and we show that in the model-based setting it…

Optimization and Control · Mathematics 2022-02-16 Benjamin Gravell , Iman Shames , Tyler Summers