English
Related papers

Related papers: Hamilton-Jacobi Deep Q-Learning for Deterministic …

200 papers

This paper introduces the Hamilton-Jacobi-Bellman Proximal Policy Optimization (HJBPPO) algorithm into reinforcement learning. The Hamilton-Jacobi-Bellman (HJB) equation is used in control theory to evaluate the optimality of the value…

Machine Learning · Computer Science 2023-02-02 Amartya Mukherjee , Jun Liu

We introduce a new numerical method to approximate the solution of a finite horizon deterministic optimal control problem. We exploit two Hamilton-Jacobi-Bellman PDE, arising by considering the dynamics in forward and backward time. This…

Optimization and Control · Mathematics 2023-04-21 Marianne Akian , Stéphane Gaubert , Shanqing Liu

An optimal control problem is considered for a stochastic differential equation containing a state-dependent regime switching, with a recursive cost functional. Due to the non-exponential discounting in the cost functional, the problem is…

Optimization and Control · Mathematics 2017-12-29 Hongwei Mei , Jiongmin Yong

We develop dynamical programming methods for the purpose of optimal control of quantum states with convex constraints and concave cost and bequest functions of the quantum state. We consider both open loop and feedback control schemes,…

Quantum Physics · Physics 2009-03-06 Viacheslav P. Belavkin , Antonio Negretti , Klaus Molmer

In this paper, we present a scalable deep learning approach to solve opinion dynamics stochastic optimal control problems with mean field term coupling in the dynamics and cost function. Our approach relies on the probabilistic…

Multiagent Systems · Computer Science 2022-04-19 Tianrong Chen , Ziyi Wang , Evangelos A. Theodorou

This paper investigates the continuous-time counterpart of the Q-function for entropy-regularized mean-field control (MFC) with controlled common noise, coined as q-function by Jia and Zhou (2023) in the single agent's model. We first show…

Optimization and Control · Mathematics 2026-05-01 Zhenjie Ren , Xiaoli Wei , Xiang Yu , Xun Yu Zhou

We address the problem of computing a control for a time-dependent nonlinear system to reach a target set in a minimal time. To solve this minimal time control problem, we introduce a hierarchy of linear semi-infinite programs, the values…

Optimization and Control · Mathematics 2023-07-04 Antoine Oustry , Matteo Tacchi

This paper proposes two algorithms for solving stochastic control problems with deep learning, with a focus on the utility maximisation problem. The first algorithm solves Markovian problems via the Hamilton Jacobi Bellman (HJB) equation.…

Computational Finance · Quantitative Finance 2024-10-15 Ashley Davey , Harry Zheng

We consider a Bolza-type optimal control problem for a dynamical system described by a fractional differential equation with the Caputo derivative of an order $\alpha \in (0, 1)$. The value of this problem is introduced as a functional in a…

Optimization and Control · Mathematics 2019-08-06 Mikhail I. Gomoyunov

This paper is concerned with a stochastic recursive optimal control problem with time delay, where the controlled system is described by a stochastic differential delayed equation (SDDE) and the cost functional is formulated as the solution…

Optimization and Control · Mathematics 2014-08-26 Jingtao Shi , Huanshui Zhang

In this paper, we are concerned with the classical solvability of a class of second-order Hamilton-Jacobi-Bellman equations (HJB equations) arising from stochastic optimal control problems with linear dynamics and uniformly convex cost…

Optimization and Control · Mathematics 2025-12-19 Jinghua Li , Zhiyong Yu

In this paper, two Q-learning (QL) methods are proposed and their convergence theories are established for addressing the model-free optimal control problem of general nonlinear continuous-time systems. By introducing the Q-function for…

Systems and Control · Computer Science 2014-10-14 Biao Luo , Derong Liu , Tingwen Huang

We propose a parallel algorithm for the numerical solution of a class of second order semi-linear equations coming from stochastic optimal control problems, by means of a dynamic domain decomposition technique. The new method is an…

Numerical Analysis · Mathematics 2016-02-11 Simone Cacace , Maurizio Falcone

Stochastic optimal control problems for Hamiltonian dynamics on graphs have wide-ranging applications in mechanics and quantum field theory, particularly in systems with graph-based structures. In this paper, we establish the existence and…

Optimization and Control · Mathematics 2025-10-01 Jianbo Cui , Tonghe Dang

This paper is concerned with finite-level quantum memory systems for retaining initial dynamic variables in the presence of external quantum noise. The system variables have an algebraic structure, similar to that of the Pauli matrices, and…

Optimization and Control · Mathematics 2026-04-01 Igor G. Vladimirov , Ian R. Petersen , Guodong Shi

We address the problem of combined stochastic and impulse control for a market maker operating in a limit order book. The problem is formulated as a Hamilton-Jacobi-Bellman quasi-variational inequality (HJBQVI). We propose an implicit…

Mathematical Finance · Quantitative Finance 2025-12-25 Alexey Meteykin

We treat infinite horizon optimal control problems by solving the associated stationary Hamilton-Jacobi-Bellman (HJB) equation numerically to compute the value function and an optimal feedback law. The dynamical systems under consideration…

Optimization and Control · Mathematics 2021-05-19 Mathias Oster , Leon Sallandt , Reinhold Schneider

Recent research reveals that deep learning is an effective way of solving high dimensional Hamilton-Jacobi-Bellman equations. The resulting feedback control law in the form of a neural network is computationally efficient for real-time…

Dynamical Systems · Mathematics 2022-10-10 Wei Kang , Qi Gong , Tenavi Nakamura-Zimmerer

Continuous-time stochastic processes underlie many natural and engineered systems. In healthcare, autonomous driving, and industrial control, direct interaction with the environment is often unsafe or impractical, motivating offline…

Machine Learning · Statistics 2025-11-14 Nicolas Hoischen , Petar Bevanda , Max Beier , Stefan Sosnowski , Boris Houska , Sandra Hirche

To sidestep the curse of dimensionality when computing solutions to Hamilton-Jacobi-Bellman partial differential equations (HJB PDE), we propose an algorithm that leverages a neural network to approximate the value function. We show that…

Machine Learning · Computer Science 2017-03-28 Frank Jiang , Glen Chou , Mo Chen , Claire J. Tomlin