English
Related papers

Related papers: Policy Iteration for Exploratory Hamilton--Jacobi-…

200 papers

This paper considers linear-quadratic control of a non-linear dynamical system subject to arbitrary cost. I show that for this class of stochastic control problems the non-linear Hamilton-Jacobi-Bellman equation can be transformed into a…

General Physics · Physics 2009-11-11 H. J. Kappen

We investigate the long time behavior of weakly dissipative semilinear Hamilton-Jacobi-Bellman (HJB) equations and the turnpike property for the corresponding stochastic control problems. To this aim, we develop a probabilistic approach…

Probability · Mathematics 2023-03-17 Giovanni Conforti

This study deals with continuous limits of interacting one-dimensional diffusive systems, arising from stochastic distortions of discrete curves with various kinds of coding representations. These systems are essentially of a…

Statistical Mechanics · Physics 2011-09-09 Guy Fayolle , Cyril Furtlehner

Unbounded stochastic control problems may lead to Hamilton-Jacobi-Bellman equations whose Hamiltonians are not always defined, especially when the diffusion term is unbounded with respect to the control. We obtain existence and uniqueness…

Analysis of PDEs · Mathematics 2008-10-09 Francesca Da Lio , Olivier Ley

We develop a discrete analogue of Hamilton-Jacobi theory in the framework of discrete Hamiltonian mechanics. The resulting discrete Hamilton-Jacobi equation is discrete only in time. We describe a discrete analogue of Jacobi's solution and…

Optimization and Control · Mathematics 2011-08-15 Tomoki Ohsawa , Anthony M. Bloch , Melvin Leok

We assume a continuous-time price impact model similar to Almgren-Chriss but with the added assumption that the price impact parameters are stochastic processes modeled as correlated scalar Markov diffusions. In this setting, we develop…

Trading and Market Microstructure · Quantitative Finance 2018-04-13 Weston Barger , Matthew Lorig

The control of relaxation-type systems of ordinary differential equations is investigated using the Hamilton-Jacobi-Bellman equation. First, we recast the model as a singularly perturbed dynamics which we embed in a family of controlled…

Optimization and Control · Mathematics 2024-04-23 Michael Herty , Hicham Kouhkouh

We propose a mesh-free policy iteration framework that combines classical dynamic programming with physics-informed neural networks (PINNs) to solve high-dimensional, nonconvex Hamilton--Jacobi--Isaacs (HJI) equations arising in stochastic…

Numerical Analysis · Mathematics 2025-07-24 Hee Jun Yang , Minjung Gim , Yeoneung Kim

This paper studies the robustness of policy iteration in the context of continuous-time infinite-horizon linear quadratic regulation (LQR) problem. It is shown that Kleinman's policy iteration algorithm is inherently robust to small…

Systems and Control · Electrical Eng. & Systems 2020-09-01 Bo Pang , Tao Bian , Zhong-Ping Jiang

In this paper we propose and analyze a method based on the Riccati transformation for solving the evolutionary Hamilton-Jacobi-Bellman equation arising from the stochastic dynamic optimal allocation problem. We show how the fully nonlinear…

Portfolio Management · Quantitative Finance 2013-07-25 Sona Kilianova , Daniel Sevcovic

In this work, we present numerical analysis for a distributed optimal control problem, with box constraint on the control, governed by a subdiffusion equation which involves a fractional derivative of order $\alpha\in(0,1)$ in time. The…

Numerical Analysis · Mathematics 2017-12-22 Bangti Jin , Buyang Li , Zhi Zhou

We propose a new numerical method for solving the Hamilton-Jacobi-Bellman quasi-variational inequality associated with the combined impulse and stochastic optimal control problem over a finite time horizon. Our method corresponds to an…

Numerical Analysis · Mathematics 2015-02-05 Masashi Ieda

The uncertainties in plant dynamics remain a challenge for nonlinear control problems. This paper develops a ternary policy iteration (TPI) algorithm for solving nonlinear robust control problems with bounded uncertainties. The controller…

Systems and Control · Electrical Eng. & Systems 2020-07-15 Jie Li , Shengbo Eben Li , Yang Guan , Jingliang Duan , Wenyu Li , Yuming Yin

We treat infinite horizon optimal control problems by solving the associated stationary Hamilton-Jacobi-Bellman (HJB) equation numerically to compute the value function and an optimal feedback law. The dynamical systems under consideration…

Optimization and Control · Mathematics 2021-05-19 Mathias Oster , Leon Sallandt , Reinhold Schneider

This work proposes a novel numerical scheme for solving the high-dimensional Hamilton-Jacobi-Bellman equation with a functional hierarchical tensor ansatz. We consider the setting of stochastic control, whereby one applies control to a…

Numerical Analysis · Mathematics 2025-07-01 Xun Tang , Nan Sheng , Lexing Ying

There has been a recent focus in reinforcement learning on addressing continuous state and action problems by optimizing parameterized policies. PI2 is a recent example of this approach. It combines a derivation from first principles of…

Machine Learning · Computer Science 2012-06-22 Freek Stulp , Olivier Sigaud

The policy gradient theorem (Sutton et al., 2000) prescribes the usage of a cumulative discounted state distribution under the target policy to approximate the gradient. Most algorithms based on this theorem, in practice, break this…

Machine Learning · Computer Science 2022-07-08 Samuele Tosatto , Andrew Patterson , Martha White , A. Rupam Mahmood

This paper considers consumption and portfolio optimization problems with recursive preferences in both infinite and finite time regions. Specially, the financial market consists of a risk-free asset and a risky asset that follows a general…

Optimization and Control · Mathematics 2024-12-30 Jian-hao Kang , Zhun Gou , Nan-jing Huang

In this work we investigate the optimal proportional reinsurance-investment strategy of an insurance company which wishes to maximize the expected exponential utility of its terminal wealth in a finite time horizon. Our goal is to extend…

Risk Management · Quantitative Finance 2019-04-04 Matteo Brachetta , Claudia Ceci

In this paper, we consider the problem of controlling a diffusion process pertaining to an opioid epidemic dynamical model with random perturbation so as to prevent it from leaving a given bounded open domain. Here, we assume that the…

Optimization and Control · Mathematics 2018-06-26 Getachew K. Befekadu , Quanyan Zhu