English
Related papers

Related papers: Data-based approximate policy iteration for nonlin…

200 papers

The Bellman equation and its continuous form, the Hamilton-Jacobi-Bellman equation, are ubiquitous in reinforcement learning and control theory. However, these equations become intractable for high-dimensional or nonlinear systems. This…

Artificial Intelligence · Computer Science 2026-05-04 Preston Rozwood , Edward Mehrez , Ludger Paehler , Wen Sun , Steven L. Brunton

We provide a data-driven framework for optimal control of a continuous-time stochastic dynamical system. The proposed framework relies on the linear operator theory involving linear Perron-Frobenius (P-F) and Koopman operators. Our first…

Optimization and Control · Mathematics 2022-02-04 Umesh Vaidya , Duvan Tellez-Castro

Controlling systems of ordinary differential equations (ODEs) is ubiquitous in science and engineering. For finding an optimal feedback controller, the value function and associated fundamental equations such as the Bellman equation and the…

Optimization and Control · Mathematics 2021-04-14 Mathias Oster , Leon Sallandt , Reinhold Schneider

In this paper, we first establish the dynamic programming principle for stochastic optimal control problems defined on compact Riemannian manifolds without boundary. Subsequently, we derive the associated Hamilton-Jacobi-Bellman (HJB)…

Optimization and Control · Mathematics 2025-07-03 Dingqian Gao , Qi Lü

Value-based reinforcement-learning algorithms provide state-of-the-art results in model-free discrete-action settings, and tend to outperform actor-critic algorithms. We argue that actor-critic algorithms are limited by their need for an…

Machine Learning · Computer Science 2019-06-13 Denis Steckelmacher , Hélène Plisnier , Diederik M. Roijers , Ann Nowé

We address finding the semi-global solutions to optimal feedback control and the Hamilton--Jacobi--Bellman (HJB) equation. Using the solution of an HJB equation, a feedback optimal control law can be implemented in real-time with minimum…

Optimization and Control · Mathematics 2016-06-17 Wei Kang , Lucas C. Wilcox

We introduce a new and efficient numerical method for multicriterion optimal control and single criterion optimal control under integral constraints. The approach is based on extending the state space to include information on a "budget"…

Optimization and Control · Mathematics 2016-01-06 Ajeet Kumar , Alexander Vladimirsky

From the Hamilton-Jacobi-Bellman equation for the value function we derive a non-linear partial differential equation for the optimal portfolio strategy (the dynamic control). The equation is general in the sense that it does not depend on…

Portfolio Management · Quantitative Finance 2013-11-20 Mads Nielsen

It is well known that time dependent Hamilton-Jacobi-Isaacs partial differential equations (HJ PDE), play an important role in analyzing continuous dynamic games and control theory problems. An important tool for such problems when they…

Optimization and Control · Mathematics 2016-05-09 Jérôme Darbon , Stanley Osher

Recent research reveals that deep learning is an effective way of solving high dimensional Hamilton-Jacobi-Bellman equations. The resulting feedback control law in the form of a neural network is computationally efficient for real-time…

Dynamical Systems · Mathematics 2022-10-10 Wei Kang , Qi Gong , Tenavi Nakamura-Zimmerer

Stochastic optimal control problems governed by delay equations with delay in the control are usually more difficult to study than the the ones when the delay appears only in the state. This is particularly true when we look at the…

Probability · Mathematics 2015-06-22 Fausto Gozzi , Federica Masiero

In this paper, we aim to solve the high dimensional stochastic optimal control problem from the view of the stochastic maximum principle via deep learning. By introducing the extended Hamiltonian system which is essentially an FBSDE with a…

Optimization and Control · Mathematics 2021-06-23 Shaolin Ji , Shige Peng , Ying Peng , Xichuan Zhang

For continuous systems modeled by dynamical equations such as ODEs and SDEs, Bellman's Principle of Optimality takes the form of the Hamilton-Jacobi-Bellman (HJB) equation, which provides the theoretical target of reinforcement learning…

Machine Learning · Computer Science 2025-10-28 Haruki Settai , Naoya Takeishi , Takehisa Yairi

The Dynamic Programming approach allows to compute a feedback control for nonlinear problems, but suffers from the curse of dimensionality. The computation of the control relies on the resolution of a nonlinear PDE, the…

Numerical Analysis · Mathematics 2019-11-14 Alessandro Alla , Luca Saluzzi

When randomness in demand affects the sales of a product, retailers use dynamic pricing strategies to maximize their profits. In this article, we formulate the pricing problem as a continuous-time stochastic optimal control problem and find…

Optimization and Control · Mathematics 2019-03-13 Asbjørn Nilsen Riseth

We revisit the linear programming approach to deterministic, continuous time, infinite horizon discounted optimal control problems. In the first part, we relax the original problem to an infinite-dimensional linear program over a measure…

Optimization and Control · Mathematics 2017-06-08 Angeliki Kamoutsi , Tobias Sutter , Peyman Mohajerin Esfahani , John Lygeros

To sidestep the curse of dimensionality when computing solutions to Hamilton-Jacobi-Bellman partial differential equations (HJB PDE), we propose an algorithm that leverages a neural network to approximate the value function. We show that…

Machine Learning · Computer Science 2017-03-28 Frank Jiang , Glen Chou , Mo Chen , Claire J. Tomlin

The goal of this article is to study fundamental mechanisms behind so-called indirect and direct data-driven control for unknown systems. Specifically, we consider policy iteration applied to the linear quadratic regulator problem. Two…

Systems and Control · Electrical Eng. & Systems 2024-04-30 Bowen Song , Andrea Iannelli

An adaptive controller is proposed and analyzed for the class of infinite-horizon optimal control problems in positive linear systems presented in (Ohlin et al., 2024b). This controller is derived from the solution of a "data-driven…

Optimization and Control · Mathematics 2025-04-22 Fethi Bencherki , Anders Rantzer

Stochastic optimal control problems governed by delay equations with delay in the control are usually more difficult to study than the the ones when the delay appears only in the state. This is particularly true when we look at the…

Probability · Mathematics 2021-03-22 F. Gozzi , F. Masiero